logoalt Hacker News

refulgentistoday at 6:14 PM1 replyview on HN

The thoughts trick was known before their paper / August.

I "independently" "invented" it for the first Anthropic reasoning models because the API required you have thoughts for each assistant message. My app lets you switch AIs within a chat, and their API used to require thinking for all messages if thinking was enabled, so I needed to get a valid thinking stub to insert.

Time has flew by for me the last 3 years, but, I'd guess it's been at least 18 months. And IMHO it wasn't very complicated to work through how to do once you were dead set on making it happen. I expect it was well-known to distillers before the paper.


Replies

7734128today at 7:03 PM

Sure, but TFA is trying to use Qwen's reaction to the thoughts as proof that they did indeed extract thoughts to train on.

My point is that any model trained after August 10 will know of those specific thoughts.

show 1 reply