logoalt Hacker News

XCSmeyesterday at 4:50 PM0 repliesview on HN

In my tests, 3.6 Flash is NOT more token efficient, so it actually ends up costing more than 3.5 Flash, even with the output price reduction.

EDIT: It less less verbose in final output though, but it reasons more.

I assume the optimization comes when you have long-running tasks with many tool calls, and by reasoning more, it reduces the number of tool calls needed.