logoalt Hacker News

jostmeytoday at 3:05 PM4 repliesview on HN

And why won’t the frontier models continue to become better? The open models are getting better but so are the frontier models. The frontier models might remain in a constant race to remain ahead


Replies

hadlocktoday at 6:11 PM

I suspect due to model distillation, their need to stay on top for IPO value is more important than absolutely crushing their competition, who will simply harvest their output and distill it for training data on a ~3 month delay. Also they are probably hitting a wall on increase in intelligence vs training time.

sweetjulytoday at 3:33 PM

I don't think it's a matter of "staying ahead"; the proprietary frontier models are better, but the trouble for them is that open weight models are good enough in increasingly many cases. This raises the floor on the frontier companies and cuts their total addressable market by commodifying the easier LLM tasks. This is really the central argument of tfa :)

show 1 reply
ForHackernewstoday at 3:18 PM

They are running out of novel, clean training data and compute. There is probably a limit to how much improvement can be squeezed out of LLMs. Recent improvements have been more about orchestration and "reasoning" loops (i.e. iteratively feeding context back through the model).

show 1 reply
HappyPanaceatoday at 3:20 PM

Diminishing returns on both intelligence and training, mostly