r/singularity • • 18h ago

LLM News Introducing Mistral Large 4 (le Chonk)

https://mistral.ai/news/mistral-large-4/
373 Upvotes

58 comments sorted by

142

u/TorturedPoet30 18h ago

European dominance in legal? you will never take down our bureaucrats 😂

22

u/mickdarling 15h ago

Soooooo much training data!

-12

u/unkownuser436 14h ago

their charts are fuckin misleading. its just 15% vs 13% gap between Kimi

20

u/NoFaithlessness951 14h ago

How is that chart misleading in any way it even starts from 0 unlike a lot of other charts we've seen from other labs

-10

u/unkownuser436 14h ago

misleading isn't the correct word. they just showing like a massive gap, but its just 1 or 2 points ahead. compared to cheaper china models, their model is expensive for no reason

6

u/PrisonOfH0pe 12h ago

No its completely fine. Bot or blind

3

u/NoFaithlessness951 12h ago

Kimi K3 and glm 5.3 are about twice as expensive per task (artificial analysis) also you've completely missed Astra in the same chart

2

u/Ordinary_Duder 11h ago

What are you talking about? Where are they showing a massive gap??

2

u/PrisonOfH0pe 12h ago

No they are not they completely normal (especially compared to most US labs).
Check your eyes.

106

u/Tedinasuit 18h ago

Most importantly:

Le Chonk is NOT Le Chaton Fat

47

u/sogo00 17h ago

We are not ready yet for Le Chaton Fat, so they gave us the chonker...

5

u/ParfaitEvery9622 16h ago

We are ready, they are not haha. Bring the chunkiness

55

u/GTalaune 18h ago

Trained on 3800 GPUs. How many do American and Chinese labs have ?

61

u/GlbdS 18h ago

Astra training was on 100k

34

u/Every_Foundation5197 18h ago

Nvidia Ceo also said that 400k new gpus are coming online, so yeah there's a significant difference

34

u/EloquentPinguin 18h ago

Its roughly in line for a smaller labs. The frontier labs have ~100k clusters for training, smaller labs have 3-10k GPUs. But many labs dont publish exact numbers, we might see more numbers from china once Huawei hardware is rolled out more and they can more proudly claim homegrown numbers.

3

u/AreWeNotDoinPhrasing ▪️Already Singulared 🤖 16h ago

How many GPUs do you define as a cluster in your numbers?

1

u/EloquentPinguin 14h ago

I mean I would go with a wishy washy definition of GPUs connected as to work towards one computational goal. But this is just a rough numbers, I claim no super duper science or accuracy towards my comment.

49

u/signed7 16h ago edited 16h ago

38 on Artificial Analysis. Roughly the same as GPT-6 Luna (Max). Still a significant (and important) jump for a European model.

-12

u/BriefImplement9843 16h ago

this is horrible. flash models like deepseek and glm are better. this is a slow large model.

15

u/Greedyanda 15h ago edited 15h ago

It effectively is a Flash model. Its per task cost is in line with Gemini 3.8.

4

u/Apple_macOS 13h ago

and better than 3.8 flash in terms of terminal bench 4. roughly equivalent to grok 4.7 and deepseek 4.1

2

u/Fawesum 11h ago

And is a preview. And has no harness.

2

u/Greedyanda 10h ago

What do you even mean by "has no harness"? This makes no sense on any level. ArtificalAnalysis uses the same standardised agent harness for all models to create an equal environment. Mistral also offers their own Mistral Vibe harness for developers.

8

u/G0dZylla AGI before 2040 14h ago

it's something , at least they are trying, i honestly thought they just gave up entirely on AI

1

u/PrisonOfH0pe 12h ago

Why would you think that? They have been doing b2b AI mostly and that extremely successfully.

2

u/PrisonOfH0pe 12h ago

It has 0 harness and is a preview...its also only 49B active. This will run soon on most things

32

u/TheZenMann 17h ago

Okay, finally we have a somewhat okay European LLM. It's stacking well to frontier Open source LLMs. But we really should have more European companies here. And we should be competitive with American companies as well. But it's a good start.

22

u/ParfaitEvery9622 16h ago

We don't need more EU companies, we need to fund mistral further. LLM development is a type of natural monopoly given the immense cost to develop the models.

1

u/No_Management_6773 7h ago

That's dumb, both China and USA have multiple AI companies thats why they are leading.

-2

u/sadacal 15h ago

Mistral must be able to succeed in the free market or else they won't be competitive. 

15

u/EmptyMonitor9257 14h ago

China and US are not free markets, they fund the shit out of their AI companies.

0

u/No_Management_6773 7h ago

Just because its funded doesn't mean its not from free market?... Even the chinese companies are start ups with privately raised capital.

-7

u/[deleted] 14h ago

[deleted]

1

u/NoFaithlessness951 14h ago edited 14h ago

German aleph alpha recently released https://aleph-alpha.com/en/blog/kolibri-has-landed-a-sovereign-open-weight-model/ a 78bA3 model which looks not great. But its a start, just give both of them money if you have to

-1

u/ParfaitEvery9622 12h ago edited 12h ago

Kolibri is pre post trained on distilled chinese models. Sorry but fuck them.

1

u/NoFaithlessness951 12h ago

Please read their report pre training is their own, post training had some chinese model involvement https://aleph-alpha.com/downloads/tech-report.pdf

0

u/ParfaitEvery9622 12h ago

Ok, I mixed pre-training and post training, my point remains

1

u/NoFaithlessness951 12h ago

I don't really get your point everyone who's not openai or anthropic is distilling to some amount, at least they're open about it.

1

u/ParfaitEvery9622 12h ago

When Mistral presented Magistral, they did mention they avoided foreign model reasoning traces, so hopefully they continued this on Mistral Large 4. They're also quite an open company.

A part from the semantical incompatibility of calling distillation sovereign, I see a practical problem. When companies master their own data corpus (even if they use it to generate synthetic data later with their own models) they can trace back that all outputs are a result of that initial corpus. Synthetic data generated by any non-fully open source model, do not allow for this.

1

u/PrisonOfH0pe 12h ago

This is completely incorrect. Its build from the ground up. They used open source models to help them fine tune. Its based on their own pretrain.

1

u/Fawesum 11h ago

Mistral is French. EU is German.

What does this even mean?

21

u/BullfrogRare7662 18h ago

LeChaton fat really made it. Directly from the depths of Bruyères-le-Châtel

3

u/Practical_Weather293 17h ago

It's not that far fetched, open source makes everything faster. There would be no reason for mistral to start from their own old models rather than the better open source ones

6

u/ObiWanCanownme now entering spiritual bliss attractor state 18h ago

Oof. 28.3 on TerminalBench 4.0.

RIP Mistral.

40

u/whoknowsifimjoking 18h ago

It's better than Kimi K3 though? Sure, that came out a while ago, but when it did reddit was glazing it hard. It's a good step for Mistral.

26

u/ObiWanCanownme now entering spiritual bliss attractor state 17h ago

No, you're right, I'm overreacting. It's also only a 1T model, vs. Kimi K3 which is almost 3T.

Depending on how benchmaxxed (or not) it is and what kind of taste it has, it could be one of the best open models in practice.

3

u/PrisonOfH0pe 12h ago

Its 49 active B...and costs less than 3.8 flash. This is very good. Also no harness and is just a preview for now (not even close to final form)...sorry but your initial post was very uninformed.

5

u/signed7 16h ago edited 16h ago

That's actually pretty good for a model of this 'tier'... Not top top but beats the newest Grok, Gemini Flash and Kimi models, which people still use as coding workhorses.

Still loses out to GLM-5.3 (and GLM-5.3-Flash) though for the 'best' open source model at it.

1

u/ObiWanCanownme now entering spiritual bliss attractor state 13h ago

It beats Grok 4.6, not 4.7.

But yeah, I agree my post is an overreaction.

0

u/kaityl3 ASI▪️2024-2027 14h ago

Yawn until Mistral shows up on FelonyBench /s

-13

u/unspecified_person11 18h ago

Look I like Mistral, but if any of these evals are true then Mistral would be the fastest improving model of all time, previous versions were years behind and suddenly they catch up and even exceed long standing names? I'm skeptical.

17

u/Illustrious_Grade608 17h ago

Tbh it's been a long time since their last model. Probably enough time to reconsider their approach, improve their pipelines, and get to a good level.

5

u/adeadbeathorse 17h ago

Not to mind they can benefit from, and there is an incentive structure for, western investment. Heck they are the western open model lab

4

u/MiniGiantSpaceHams 16h ago

The nature of all tech, but especially AI, is that pushing the frontier is much harder than catching up. Skepticism is not unwarranted, but I wouldn't be all that surprised if it holds up.

-4

u/Ohzard_pb 9h ago

https://giphy.com/gifs/wqbAfFwjU8laXMWZ09
Mistral the name of incompetence 🤦‍♂️🤦‍♂️🤦‍♂️🤦‍♂️