logoalt Hacker News

valleyeryesterday at 7:17 PM1 replyview on HN

During training, certain tokens are more likely to lead to a lower loss function value, which is how you "win" the game of LLM output.


Replies

mwkaufmayesterday at 7:20 PM

So, next-token predictors

show 1 reply