logoalt Hacker News

florianherrengttoday at 4:18 PM2 repliesview on HN

This paper puts words to something I’ve noticed repeatedly with LLMs, particularly Qwen3.6. When I read its reasoning, it appears to recognise the mistake and then carry on as if it hadn’t noticed it at all.

> models often determine their answers based on implicit biases tied to question templates, then construct reasoning chains to justify their predetermined conclusions > its reasoning was correct right until the final step (Yes/No answer)


Replies

Georgelementaltoday at 4:48 PM

Natural intelligences do this too

show 3 replies
paimapitoday at 5:26 PM

[dead]