logoalt Hacker News

Mistral Large 4

1617 points • by Philpax • yesterday at 1:15 PM • 972 comments • view on HN

https://docs.mistral.ai/models/mistral-large-4-0


Comments

Kim_Bruning • yesterday at 2:14 PM

I tried a quick abbreviated kimbench on their playground before bothering to do the whole thing.

Maybe I didn't really select mistral 4? Either way, failed completely on question 1 and the next 2 questions were completely off base too. I didn't bother to finish.

Not suitable for my purposes I don't think.

andhuman • yesterday at 2:56 PM

At the end of the blog post we get this nugget.

> The pace of progress from here will be fast. Stay tuned.

brendong • yesterday at 4:14 PM

Is the reason for the massive gains in certain benchmarks due to distillation from the other lead models hence the slightly "under" pattern seen in the comparison charts?

valzam • yesterday at 2:21 PM

What experience have people had with Mistral models for cyber research? are they as constrainted as Anthropic models? I use claude day2day but have the need for a model with fewer guardrails to pentest our own APIs.

➕ show 1 reply
aennassiri • yesterday at 2:24 PM

Excited to see this! Nice that they are saying this is just a first step.

Give them more compute!

XCSme • yesterday at 4:25 PM

I tried testing it, but reasoning effort parameter doesn't seem to be there and output is sort of broken because it reasons directly in the output tokens...

ktosobcy • yesterday at 4:36 PM

Awesome!

All things considered I'm more inclained to pay for EU-based AI in the end (supporting local company and most likely being more aligned with EU regulations…)

rarisma • yesterday at 1:49 PM

le chaton fat is real, my life is complete. Benches look crazy good for 1T.

scirob • yesterday at 5:19 PM

no hugging face link :( ... but hey its on openrouter yay

Here are my results

https://dach.peerbench.ai/compare?models=mistralai%2Fmistral...

Looks like a bit better than the recent Kolibri-1 but still below Qwen3.8 27B

jrflo • yesterday at 2:19 PM

Seems like they've finally made a genuinely competitive model since the original LLM craze, congrats to them! Glad to see some diversification in open weights providers.

Roark66 • yesterday at 1:32 PM

This sure is nice, but I've had less than satisfactory results with GLM5.3. I'd like Mistral to compete with Qwen3.8-Flash-Next a 120B class model that IMO is the first model that I can use for serious coding while running it locally.

I estimate it's coding ability on par with opus 4.6 (but opus definitely beats it on factual knowledge) Still it's a genuinely useful model, when everything else except Anthropic's models (and for only 3 weeks after it came out Google's Gemini 3 pro, before it got merged) are not to me.

I'd live to have one like that but EU made.

blauditore • yesterday at 4:50 PM

What's up with the name? It reminds me of my teenage self trying to speak in funny memes.

carodgers • yesterday at 4:28 PM

Without exaggeration, given a choice between models, I would pay for Mistral's model over Anthropic's based on the name alone, completely ignoring features or other technical considerations. The name is playful and is such a refreshing contrast to Anthropic's (and OpenAI's) doomsaying, scaremongering, and god-posturing.

armaghanraza • yesterday at 5:15 PM

Does it have the ability to capture the market like OpenAi or Anthropic ? My point is, Regular/Average users of AI do not really care about benchmarking. Marketing decides which company makes it to the phones or PCs.

➕ show 1 reply
sixhobbits • yesterday at 2:04 PM

if it's not available yet why have a 'try it today' header at all?

> "Try it today" > > There is still more to come. As we work toward releasing the weights, we will share further details on the model architecture, additional benchmarks, and our post-training methodology.

➕ show 1 reply
staticman2 • yesterday at 1:24 PM

Since the Chinese companies publish their research it would have been odd if Mistral didn't start catching up.

➕ show 3 replies
whatever1 • yesterday at 6:27 PM

I love the name! Teasing the ones making fun of them.

timcobb • yesterday at 2:58 PM

> Trained from scratch

How are they training without pirating the Z library corpus and all that?

➕ show 1 reply
chriskanan • yesterday at 5:09 PM

I really hate this open weight but closed science approach. These companies just take from academics and all the Chinese companies that are doing good science, but without understanding the recipe it makes it hard to know where the failure points will be until your agent accidentally commits a crime.

Mistral doesn't publish the science.

➕ show 1 reply
ThouYS • yesterday at 1:49 PM

Glad to see progress, despite the ever-increasing sabotage by the EU bureaucrats

drbscl • yesterday at 2:38 PM

So about 2 or 3 generations behind, just like they were a year ago?

tosh • yesterday at 1:40 PM

sorting the charts like that gives off weird vibes

https://mistral.ai/news/mistral-large-4/

➕ show 1 reply
Nux • yesterday at 3:54 PM

Number one in Sovereign AI. Join our Discord.

AM1010101 • yesterday at 5:18 PM

Half price on open router right now

tdubey • yesterday at 1:30 PM

Is there consensus on if this was https://openrouter.ai/stealth/space-bunny-alpha ?

➕ show 2 replies
tosh • yesterday at 1:36 PM

sorting the charts like that gives off weird vibes

➕ show 2 replies
sourcecodeplz • yesterday at 6:25 PM

at 200M tokens for the full AI suite run its not token efficient at all

EDM115 • yesterday at 3:03 PM

We actually got Le Chaton Fat before GTA 6

bloodmoon • yesterday at 8:37 PM

dont take it personally, i just dont understand why to release a model that is not showing new strong capabilities, why would anybody use this model and not Claude Opus.

➕ show 1 reply
maz1b • yesterday at 2:24 PM

I'm glad they're keeping at it!

igleria • yesterday at 2:32 PM

I thought lechonk motto was just a meme!

4rtem • yesterday at 1:32 PM

Previous one is barely in top 50 on arena.ai

theturtletalks • yesterday at 2:00 PM

A bit disappointing to see it still lagging behind Chinese open models. Those Chinese models are pushing proprietary models to raise the bar, but we need equally strong non-Chinese open models to challenge the Chinese ones in turn.

ridth • yesterday at 9:50 PM

Why is distillation weird?

pietz • yesterday at 2:04 PM

I mean no disrespect but these are terrible numbers or am I missing something? It seems like Mistral continues to only be relevant for people that want a model trained in Europe. Too bad.

➕ show 1 reply
redanddead • yesterday at 2:15 PM

better than K3 and DS4, cool

maxdo • yesterday at 1:32 PM

Not bad only two major releases behind top tier. Edit : checked its rather 3 generations behind . Oh well

saberience • yesterday at 1:32 PM

Looks like it's about a year behind still. i.e. its intelligence is behind models from roughly a year ago.

https://www.vals.ai/benchmarks/vals_index

ofirg • yesterday at 1:39 PM

where does sit on the pareto distribution compered to Le Chaton Fat?

gabe-santana • yesterday at 3:36 PM

Amazing! just tested

theanimeshs • yesterday at 4:00 PM

impressive release this time by Mistral. bullish.

alpineman • yesterday at 1:59 PM

>> Unofficially ML4, very officially: le Chonk

Honestly just nice to see a leader in this space not take themselves so seriously.

LoganDark • yesterday at 5:41 PM

1T parameters -- ugh, open models keep getting bigger and bigger! Running them at home is getting ever more unattainable, especially for those of us with bandwidth-poor hardware like Apple silicon -- please continue releasing smaller models, too!

htrp • yesterday at 2:10 PM

europe finally getting into the race here.

TokiBot • yesterday at 6:17 PM

Where can it be tested?

glerk • yesterday at 5:11 PM

Massive fumble not to call it “le chaton fat”.

0xbadcafebee • yesterday at 5:03 PM

Too bad this got marked as a dupe, as it actually has benchmark info unlike the other page which is just docs.

The weird thing is how worse they are at things like coding than other open weights. You'd expect them to at least distill coding from other open weights to match them.

jvwww • yesterday at 4:37 PM

Pretty impressive. I genuinely wonder how Mistral hires talent when their salaries are so terrible. Guess there aren't many better places to work in Europe.

🔗 View 27 more comments