‹ PublicAI Index
The LLM benchmark aggregator.
Granite 3.2 2B Instruct
IBM · 2B
Strongest in Toxicity avoidance (#15 of 272), weakest in Fairness (#212 of 299). Above par in 2 of 6 scopes. Among the models it meets almost everywhere, it finishes behind Claude Sonnet 5 and Claude 3.5 Sonnet and ahead of Gemma 4 26B A4B and Granite 3.0 2B Instruct.
§ 1 · Profile
What it is good at
Bars run from 50 — the average of the models each source lists — so right of the line is above par. The middle column is the gap to whoever leads that scope.
Safety49.5−12#200/3361/3
Toxicity avoidance58.7−0.3#15/2721/1
Jailbreak resistance55.9−9#103/2721/1
Harm refusal48.5−13.1#178/2991/2
Secure code42.2−24.3#206/2741/1
Fairness44.2−25.8#212/2991/2
§ 2 · Head to head
What it beats, and what beats it
The same models turn up scope after scope. Counted once: where both were placed, who finished higher. Bars run right for Granite 3.2 2B Instruct, left for the other.
§ 3 · Sources
Where the numbers come from
1 publication, 6 figures. Every one links to the page it was read from.
Enkrypt · Jailbreak risk 7.8%Enkrypt · Harmful content risk 45%Enkrypt · CBRN risk 12.5%Enkrypt · Toxicity risk 0.2%Enkrypt · Bias risk 86.8%Enkrypt · Insecure code risk 49.3%
Badge
[](https://publicai.io/model-index/m/granite-3-2-2b-instruct)