‹ PublicAI Index
The LLM benchmark aggregator.
Jamba 1.5 Mini
AI21
Strongest in Secure code (#48 of 274), weakest in Human preference (#257 of 342). Above par in 2 of 12 scopes. Among the models it meets almost everywhere, it finishes behind Claude Opus 5.5 and Claude Opus 5 and ahead of Qwen3 30B A3B Instruct and GPT-3.5 Turbo.
§ 1 · Profile
What it is good at
Bars run from 50 — the average of the models each source lists — so right of the line is above par. The middle column is the gap to whoever leads that scope.
Professional36.1−25.5#167/1681/1
Legal41.1−26#128/1511/1
Medical34.9−31.9#130/1401/1
Finance30−32.9#149/1511/1
Safety47.2−14.3#248/3371/3
Secure code62.1−4.4#48/2741/1
Jailbreak resistance50.2−14.7#180/2721/1
Fairness43.5−26.5#234/3001/2
Toxicity avoidance39.6−19.4#243/2721/1
Harm refusal44.6−17#253/3001/2
Human preference41.6−25.9#257/3421/1
Human preference41.6−25.9#257/3421/1
§ 2 · Head to head
What it beats, and what beats it
The same models turn up scope after scope. Counted once: where both were placed, who finished higher. Bars run right for Jamba 1.5 Mini, left for the other.
§ 3 · Sources
Where the numbers come from
3 publications, 11 figures. Every one links to the page it was read from.
LMArena Text 1240
Vals · LegalBench 66.62%Vals · CorpFin 33.88%Vals · TaxEval 41.86%Vals · MedQA 55.18%
Enkrypt · Jailbreak risk 12.6%Enkrypt · Harmful content risk 61.7%Enkrypt · CBRN risk 14%Enkrypt · Toxicity risk 13.6%Enkrypt · Bias risk 87.9%Enkrypt · Insecure code risk 8.9%
Badge
[](https://publicai.io/model-index/m/jamba-1-5-mini)