logoalt Hacker News

Wowfunhappyyesterday at 5:44 PM1 replyview on HN

I know that's true for Qwen but I don't think most models work that way?


Replies

wren6991yesterday at 6:05 PM

OpenAI models also work this way, as evidenced by full cache blowout when changing reasoning level. Every single open-weight model I've seen also works this way (your "reasoning_effort" argument just changes a small section of the system prompt in the chat template). I would have to see some evidence to believe Anthropic were doing anything different.