logoalt Hacker News

ac-cianoyesterday at 3:22 PM1 replyview on HN

Yeah it's exactly what I've seen as well. Ordewell is close to your high end case: a frontier model builds the plan once, then it's fixed, the executing model can't reinterpret it. Won't get you to 4B, but should help a small model that only has to execute, not plan and execute at once.


Replies

hedgehogyesterday at 3:37 PM

Have you quantified the performance on any particular benchmark?

show 1 reply