‹ PublicAI Index
The LLM benchmark aggregator.
Mistral 7B Instruct V0.2
Mistral · 7B
Strongest in Toxicity avoidance (#78 of 272), weakest in Human preference (#305 of 342). Above par in 2 of 8 scopes. Among the models it meets almost everywhere, it finishes behind Claude Opus 4.7 and Claude Opus 4.8 and ahead of Command R and Grok 4.1 Fast.
§ 1 · Profile
What it is good at
Bars run from 50 — the average of the models each source lists — so right of the line is above par. The middle column is the gap to whoever leads that scope.
Safety46.8−14.7#259/3371/3
Toxicity avoidance56−3#78/2721/1
Jailbreak resistance55.5−9.4#110/2721/1
Fairness44.2−25.8#210/3001/2
Harm refusal45.9−15.7#227/3001/2
Secure code33.2−33.3#240/2741/1
Human preference32.8−34.7#305/3421/1
Human preference32.8−34.7#305/3421/1
§ 2 · Head to head
What it beats, and what beats it
The same models turn up scope after scope. Counted once: where both were placed, who finished higher. Bars run right for Mistral 7B Instruct V0.2, left for the other.
§ 3 · Sources
Where the numbers come from
2 publications, 7 figures. Every one links to the page it was read from.
LMArena Text 1149
Enkrypt · Jailbreak risk 8.1%Enkrypt · Harmful content risk 60%Enkrypt · CBRN risk 11.3%Enkrypt · Toxicity risk 2.1%Enkrypt · Bias risk 86.8%Enkrypt · Insecure code risk 67.6%
Badge
[](https://publicai.io/model-index/m/mistral-7b-instruct-v0-2)