‹ PublicAI Index
The LLM benchmark aggregator.
Laguna M.1
Other
Strongest in Academic knowledge (#112 of 123), weakest in Professional (#163 of 168). Above par in 0 of 10 scopes. Among the models it meets almost everywhere, it finishes behind Claude Fable 5.1 and Claude Fable 5 and ahead of Command R Plus and Jamba 1.6 Mini.
§ 1 · Profile
What it is good at
Bars run from 50 — the average of the models each source lists — so right of the line is above par. The middle column is the gap to whoever leads that scope.
Knowledge37.1−26.4#128/1381/2
Academic knowledge34.5−29.4#112/1231/1
Coding41.8−27.3#144/1651/5
Agentic coding40.6−28.1#141/1571/4
Reasoning38.8−28.1#157/1781/4
Science26.1−37.4#122/1221/1
Professional37−24.6#163/1681/1
Legal39.2−27.9#133/1511/1
Medical33.5−33.3#133/1401/1
Finance34.9−28#145/1511/1
§ 2 · Head to head
What it beats, and what beats it
The same models turn up scope after scope. Counted once: where both were placed, who finished higher. Bars run right for Laguna M.1, left for the other.
§ 3 · Sources
Where the numbers come from
1 publication, 12 figures. Every one links to the page it was read from.
Vals · Legal Research Bench 2.4%Vals · LegalBench 75.14%Vals · Harvey Legal Agent Benchmark 0%Vals · Finance Agent 25.03%Vals · CorpFin 58.16%Vals · TaxEval 1.64%Vals · MedCode 23.11%Vals · MedScribe 65.91%Vals · SWE-bench Verified 57.6% (Mini-SWE-agent)Vals · Vibe Code Bench 11.04% (OpenHands)Vals · GPQA Diamond 27.02%Vals · MMLU Pro 68.84%
Badge
[](https://publicai.io/model-index/m/laguna-m-1)