‹ PublicAI Index
The LLM benchmark aggregator.
Magistral Small 1.2
Mistral
Strongest in Mathematics (#63 of 140, on 1 of its 2 boards), weakest in Agents (#222 of 268). Above par in 1 of 11 scopes. Among the models it meets almost everywhere, it finishes behind Claude Opus 5 and Claude Fable 5 and ahead of Mistral Small 3.1 and Mercury 2.5.
§ 1 · Profile
What it is good at
Bars run from 50 — the average of the models each source lists — so right of the line is above par. The middle column is the gap to whoever leads that scope.
Reasoning47.7−19.2#102/1781/4
Mathematics53.8−11.9#63/1401/2
Science37.6−25.9#102/1221/1
Knowledge30.1−33.4#135/1381/2
Academic knowledge26.1−37.8#120/1231/1
Professional42.2−19.4#147/1681/1
Medical49.4−17.4#83/1401/1
Finance41.1−21.8#125/1511/1
Legal35−32.1#150/1511/1
Agents41.7−26.4#222/2681/5
Knowledge work37.6−36#142/1781/2
§ 2 · Head to head
What it beats, and what beats it
The same models turn up scope after scope. Counted once: where both were placed, who finished higher. Bars run right for Magistral Small 1.2, left for the other.
§ 3 · Sources
Where the numbers come from
2 publications, 9 figures. Every one links to the page it was read from.
GDPval-AA -10
Vals · LegalBench 40.02%Vals · CorpFin 44.02%Vals · TaxEval 60.3%Vals · MortgageTax 62.12%Vals · MedQA 82.36%Vals · GPQA Diamond 58.33%Vals · MMLU Pro 62.13%Vals · AIME 80.68%
Badge
[](https://publicai.io/model-index/m/magistral-small-1-2)