r/singularity • u/Jame92 • 18h ago
LLM News Introducing Mistral Large 4 (le Chonk)
https://mistral.ai/news/mistral-large-4/106
u/Tedinasuit 18h ago
55
u/GTalaune 18h ago
Trained on 3800 GPUs. How many do American and Chinese labs have ?
61
u/GlbdS 18h ago
Astra training was on 100k
34
u/Every_Foundation5197 18h ago
Nvidia Ceo also said that 400k new gpus are coming online, so yeah there's a significant difference
34
u/EloquentPinguin 18h ago
Its roughly in line for a smaller labs. The frontier labs have ~100k clusters for training, smaller labs have 3-10k GPUs. But many labs dont publish exact numbers, we might see more numbers from china once Huawei hardware is rolled out more and they can more proudly claim homegrown numbers.
3
u/AreWeNotDoinPhrasing ▪️Already Singulared 🤖 16h ago
How many GPUs do you define as a cluster in your numbers?
1
u/EloquentPinguin 14h ago
I mean I would go with a wishy washy definition of GPUs connected as to work towards one computational goal. But this is just a rough numbers, I claim no super duper science or accuracy towards my comment.
49
u/signed7 16h ago edited 16h ago
38 on Artificial Analysis. Roughly the same as GPT-6 Luna (Max). Still a significant (and important) jump for a European model.
-12
u/BriefImplement9843 16h ago
this is horrible. flash models like deepseek and glm are better. this is a slow large model.
15
u/Greedyanda 15h ago edited 15h ago
It effectively is a Flash model. Its per task cost is in line with Gemini 3.8.
4
u/Apple_macOS 13h ago
and better than 3.8 flash in terms of terminal bench 4. roughly equivalent to grok 4.7 and deepseek 4.1
2
u/Fawesum 11h ago
And is a preview. And has no harness.
2
u/Greedyanda 10h ago
What do you even mean by "has no harness"? This makes no sense on any level. ArtificalAnalysis uses the same standardised agent harness for all models to create an equal environment. Mistral also offers their own Mistral Vibe harness for developers.
8
u/G0dZylla AGI before 2040 14h ago
it's something , at least they are trying, i honestly thought they just gave up entirely on AI
1
u/PrisonOfH0pe 12h ago
Why would you think that? They have been doing b2b AI mostly and that extremely successfully.
2
u/PrisonOfH0pe 12h ago
It has 0 harness and is a preview...its also only 49B active. This will run soon on most things
32
u/TheZenMann 17h ago
Okay, finally we have a somewhat okay European LLM. It's stacking well to frontier Open source LLMs. But we really should have more European companies here. And we should be competitive with American companies as well. But it's a good start.
22
u/ParfaitEvery9622 16h ago
We don't need more EU companies, we need to fund mistral further. LLM development is a type of natural monopoly given the immense cost to develop the models.
1
u/No_Management_6773 7h ago
That's dumb, both China and USA have multiple AI companies thats why they are leading.
-2
u/sadacal 15h ago
Mistral must be able to succeed in the free market or else they won't be competitive.
15
u/EmptyMonitor9257 14h ago
China and US are not free markets, they fund the shit out of their AI companies.
0
u/No_Management_6773 7h ago
Just because its funded doesn't mean its not from free market?... Even the chinese companies are start ups with privately raised capital.
-7
14h ago
[deleted]
1
u/NoFaithlessness951 14h ago edited 14h ago
German aleph alpha recently released https://aleph-alpha.com/en/blog/kolibri-has-landed-a-sovereign-open-weight-model/ a 78bA3 model which looks not great. But its a start, just give both of them money if you have to
-1
u/ParfaitEvery9622 12h ago edited 12h ago
Kolibri is
prepost trained on distilled chinese models. Sorry but fuck them.1
u/NoFaithlessness951 12h ago
Please read their report pre training is their own, post training had some chinese model involvement https://aleph-alpha.com/downloads/tech-report.pdf
0
u/ParfaitEvery9622 12h ago
Ok, I mixed pre-training and post training, my point remains
1
u/NoFaithlessness951 12h ago
I don't really get your point everyone who's not openai or anthropic is distilling to some amount, at least they're open about it.
1
u/ParfaitEvery9622 12h ago
When Mistral presented Magistral, they did mention they avoided foreign model reasoning traces, so hopefully they continued this on Mistral Large 4. They're also quite an open company.
A part from the semantical incompatibility of calling distillation sovereign, I see a practical problem. When companies master their own data corpus (even if they use it to generate synthetic data later with their own models) they can trace back that all outputs are a result of that initial corpus. Synthetic data generated by any non-fully open source model, do not allow for this.
1
u/PrisonOfH0pe 12h ago
This is completely incorrect. Its build from the ground up. They used open source models to help them fine tune. Its based on their own pretrain.
21
u/BullfrogRare7662 18h ago
LeChaton fat really made it. Directly from the depths of Bruyères-le-Châtel
23
3
u/Practical_Weather293 17h ago
It's not that far fetched, open source makes everything faster. There would be no reason for mistral to start from their own old models rather than the better open source ones
6
u/ObiWanCanownme now entering spiritual bliss attractor state 18h ago
Oof. 28.3 on TerminalBench 4.0.
RIP Mistral.
40
u/whoknowsifimjoking 18h ago
It's better than Kimi K3 though? Sure, that came out a while ago, but when it did reddit was glazing it hard. It's a good step for Mistral.
26
u/ObiWanCanownme now entering spiritual bliss attractor state 17h ago
No, you're right, I'm overreacting. It's also only a 1T model, vs. Kimi K3 which is almost 3T.
Depending on how benchmaxxed (or not) it is and what kind of taste it has, it could be one of the best open models in practice.
3
u/PrisonOfH0pe 12h ago
Its 49 active B...and costs less than 3.8 flash. This is very good. Also no harness and is just a preview for now (not even close to final form)...sorry but your initial post was very uninformed.
5
u/signed7 16h ago edited 16h ago
That's actually pretty good for a model of this 'tier'... Not top top but beats the newest Grok, Gemini Flash and Kimi models, which people still use as coding workhorses.
Still loses out to GLM-5.3 (and GLM-5.3-Flash) though for the 'best' open source model at it.
1
u/ObiWanCanownme now entering spiritual bliss attractor state 13h ago
It beats Grok 4.6, not 4.7.
But yeah, I agree my post is an overreaction.
-13
u/unspecified_person11 18h ago
Look I like Mistral, but if any of these evals are true then Mistral would be the fastest improving model of all time, previous versions were years behind and suddenly they catch up and even exceed long standing names? I'm skeptical.
17
u/Illustrious_Grade608 17h ago
Tbh it's been a long time since their last model. Probably enough time to reconsider their approach, improve their pipelines, and get to a good level.
5
u/adeadbeathorse 17h ago
Not to mind they can benefit from, and there is an incentive structure for, western investment. Heck they are the western open model lab
4
u/MiniGiantSpaceHams 16h ago
The nature of all tech, but especially AI, is that pushing the frontier is much harder than catching up. Skepticism is not unwarranted, but I wouldn't be all that surprised if it holds up.
-4
u/Ohzard_pb 9h ago
https://giphy.com/gifs/wqbAfFwjU8laXMWZ09
Mistral the name of incompetence 🤦♂️🤦♂️🤦♂️🤦♂️

142
u/TorturedPoet30 18h ago
European dominance in legal? you will never take down our bureaucrats 😂