logoalt Hacker News

unbrice • yesterday at 10:42 PM • 1 reply • view on HN

> the problem is newer models are never trained from scratch

Training a new base model from scratch happens every so often. Closed labs do not publish which models are new base models but as a rule of thumb major release numbers are an indication (with some exceptions).


Replies

bottlepalm • yesterday at 11:55 PM

If the training data is the same, the training algorithms are the same, the RLHF is the same, and the rest of the process is the same, then it's not really from scratch, or not from scratch in a way that results in an 'out of family' model. I doubt any company would take that risk. You always build on and use what works and go from there.