logoalt Hacker News

cube2222today at 12:26 PM2 repliesview on HN

Quickly reading the article, one notable limitation seems to be that these checkpoints are 512-1024 tokens context size models, while Jev is seemingly 32k.

That's a pretty big limitation, I would argue, unless I'm misunderstanding and it can be worked around easily somehow? I'm surprised it isn't surfaced more prominently in the comparison.


Replies

bjt12345today at 12:50 PM

Jev has 64k total token request budget and I do wonder how it will handle highly specialised inputs.

This Jev waitlist that Typesafe AI are utilising is surely going to raise questions pretty soon - it's hard to sell this to bosses when it looks like a pop-up restaurant

show 3 replies
druskaciktoday at 3:44 PM

Yeah, it's weird, considering ModernBERT, which the Laya models are based on, supports 8192 context window.