The problem with this is obviously that the only GPT 5.5 thoughts that we have access to are from stolen thought.
Qwen 3.8 0902 was trained after the release of the paper on August 10, so it should have seen those specific thoughts.
The thoughts trick was known before their paper / August.
I "independently" "invented" it for the first Anthropic reasoning models because the API required you have thoughts for each assistant message. My app lets you switch AIs within a chat, and their API used to require thinking for all messages if thinking was enabled, so I needed to get a valid thinking stub to insert.
Time has flew by for me the last 3 years, but, I'd guess it's been at least 18 months. And IMHO it wasn't very complicated to work through how to do once you were dead set on making it happen. I expect it was well-known to distillers before the paper.
seems like only the companies in question could run this sort analysis long-term; since they have full access to their CoTs not in public datasets.