logoalt Hacker News

dandieptoday at 6:33 AM4 repliesview on HN

Why are LLMs producing this super tense text more often now? Is it because they are being optimized to use fewer tokens?


Replies

vanviegentoday at 6:42 AM

I think they're overcorrecting for being too long-winded in previous generations (and still, in some cases). I guess this is a hard balance to get right.

Davidzhengtoday at 6:58 AM

I'm pretty sure the other answers are wrong and it's a side effect of RL (see thinking machines post about inkling training). It's also exacerbated in fable and sol--I think it's token efficiency effect--bc it's about to reason with fewer tokens the density of the token information goes up.

conceptiontoday at 6:36 AM

I figure they are being optimized to write code/functions and not prose so all text is getting more code like.

zahlmantoday at 6:52 AM

ChatGPT is still plenty verbose by default IMX.