logoalt Hacker News

Mistral Large 4

872 points • by Philpax • today at 1:15 PM • 521 comments • view on HN

Comments

goatley • today at 2:36 PM

Finally, a model small enough to self-host on my 2012 MacBook Air if I don't mind my house reaching room temperature in 2 seconds.

➕ show 1 reply
tdubey • today at 1:30 PM

Is there consensus on if this was https://openrouter.ai/stealth/space-bunny-alpha ?

➕ show 2 replies
4rtem • today at 1:32 PM

Previous one is barely in top 50 on arena.ai

staticman2 • today at 1:24 PM

Since the Chinese companies publish their research it would have been odd if Mistral didn't start catching up.

➕ show 3 replies
ofirg • today at 1:39 PM

where does sit on the pareto distribution compered to Le Chaton Fat?

alpineman • today at 1:59 PM

>> Unofficially ML4, very officially: le Chonk

Honestly just nice to see a leader in this space not take themselves so seriously.

htrp • today at 2:10 PM

europe finally getting into the race here.

pietz • today at 2:04 PM

I mean no disrespect but these are terrible numbers or am I missing something? It seems like Mistral continues to only be relevant for people that want a model trained in Europe. Too bad.

➕ show 1 reply
scrubmunch • today at 1:30 PM

wowza le models a heckin chonker

petesergeant • today at 1:49 PM

Anyone have any indication when I can get my hands on a developer plan for this?

saberience • today at 1:32 PM

Looks like it's about a year behind still. i.e. its intelligence is behind models from roughly a year ago.

https://www.vals.ai/benchmarks/vals_index

crimsoneer • today at 1:22 PM

Woah, this seems like a big deal (assuming the benchmarks are as good as claimed)?

Mistral slightly proving me wrong (and I'm not mad).

retinaros • today at 3:01 PM

quick question why put GLM 5.3 at 61 while a quick check on DeepSWE 1.1 puts it at 69?

also they forgot muse spark at 75% while claiming they were outshining all US models?

bdcravens • today at 1:54 PM

Can we consolidate the posts? Currently there's 3 on the front page, basically all pointing to Mistral's messaging in different places.

nicolamanzini • today at 4:42 PM

[dead]

nicolamanzini • today at 4:09 PM

[dead]

lucagiftzek • today at 4:27 PM

[dead]

sparrowidle • today at 1:20 PM

[flagged]

msavara • today at 1:23 PM

[flagged]

delillos • today at 1:21 PM

wow, another large language model from another company. groundbreaking.

➕ show 5 replies
erichocean • today at 1:29 PM

Does Mistral ever advance the state of the art on any dimension?

And if not, why do they exist?

Update: The number of people advocating not innovating is wild. There is no reason why Mistral cannot innovate in ML, they explicitly choose not to. My point is that, given that choice, they should spend their GPU hours differently.

"Sovereign AI" is a joke, there is no substantive difference between a post-trained open weight model from an American or Chinese company and what Mistral is doing today, beyond spending 80% of their GPU hours reproducing a last-gen model's pretraining.

➕ show 6 replies
maxdo • today at 1:32 PM

Not bad only two major releases behind top tier. Edit : checked its rather 3 generations behind . Oh well