logoalt Hacker News

aesthesiatoday at 2:16 AM0 repliesview on HN

Fun, though as hinted at the end, the point of LLM "truth" probes is to measure the model's internal judgment of truthfulness. There's no reason this judgment, even if measured with 100% accuracy, couldn't be mistaken or logically inconsistent.