logoalt Hacker News

bottlepalm • yesterday at 9:32 PM • 1 reply • view on HN

The problem is newer models are never trained from scratch, they generally just layer on more training data and use the same tools/methods for RLHF. OpenAI, Anthropic, xAI models all have a feel to them that carries over from one generation to the next.

Point is, if Gemini is flawed then there's a very good chance that it's still deeply flawed today, and getting smarter at the same time - that is a very bad combination.


Replies

unbrice • yesterday at 10:42 PM

> the problem is newer models are never trained from scratch

Training a new base model from scratch happens every so often. Closed labs do not publish which models are new base models but as a rule of thumb major release numbers are an indication (with some exceptions).

➕ show 1 reply