logoalt Hacker News

simonw • today at 7:11 PM • 1 reply • view on HN

It's the output token limit, which has been 128,000 for Claude models for quite a while note


Replies

croemer • today at 7:31 PM

Pretty crazy that the model doesn't know that it needs to stop before it hits 128k output tokens. I guess it has no sense of how many tokens in it is? Wouldn't this be possible to work into the architecture?

➕ show 1 reply