logoalt Hacker News

arjietoday at 5:50 PM1 replyview on HN

Is it actually entirely a prompt-based information? I’d assume that some of it is the harness part of the agent setting reasoning token budget and compacting reasoning etc.

In that case, the agent will respond incorrectly because it has no visibility into what reasoning mode it’s in.


Replies

willy_ktoday at 6:03 PM

IIRC responding to effort level settings appropriately is part of the (post)-training. In that case it could be considered another instance of the Bitter Lesson. Uplifting.