‹ PublicAI Index
The LLM benchmark aggregator.
Jamba 1.5 Large
AI21
Strongest in Legal (#94 of 151), weakest in Safety (#265 of 337). Above par in 1 of 12 scopes. Among the models it meets almost everywhere, it finishes behind Claude Opus 5.5 and Claude Opus 4.7 and ahead of Gemma 2 9B It and Llama 3.1 8B Instruct.
§ 1 · Profile
What it is good at
Bars run from 50 — the average of the models each source lists — so right of the line is above par. The middle column is the gap to whoever leads that scope.
Professional41−20.6#154/1681/1
Legal46.9−20.2#94/1511/1
Medical41.8−25#121/1401/1
Finance34.3−28.6#146/1511/1
Human preference46.3−21.2#229/3421/1
Human preference46.3−21.2#229/3421/1
Safety46.5−15#265/3371/3
Jailbreak resistance52.8−12.1#146/2721/1
Harm refusal47.6−14#198/3001/2
Fairness44.7−25.3#201/3001/2
Secure code40.5−26#220/2741/1
Toxicity avoidance44.2−14.8#234/2721/1
§ 2 · Head to head
What it beats, and what beats it
The same models turn up scope after scope. Counted once: where both were placed, who finished higher. Bars run right for Jamba 1.5 Large, left for the other.
§ 3 · Sources
Where the numbers come from
3 publications, 11 figures. Every one links to the page it was read from.
LMArena Text 1289
Vals · LegalBench 74.16%Vals · CorpFin 39.43%Vals · TaxEval 58.18%Vals · MedQA 68.11%
Enkrypt · Jailbreak risk 10.4%Enkrypt · Harmful content risk 56.1%Enkrypt · CBRN risk 9%Enkrypt · Toxicity risk 10.4%Enkrypt · Bias risk 86.1%Enkrypt · Insecure code risk 52.9%
Badge
[](https://publicai.io/model-index/m/jamba-1-5-large)