‹ PublicAI Index
The LLM benchmark aggregator.
DBRX Instruct
Databricks
Strongest in Safe-prompt compliance (#80 of 82), weakest in Safety (#332 of 337). Above par in 1 of 9 scopes. Among the models it meets almost everywhere, it finishes behind GPT-5.1 and GPT-5 and ahead of Gemma 7B It and Gemma 2B It.
§ 1 · Profile
What it is good at
Bars run from 50 — the average of the models each source lists — so right of the line is above par. The middle column is the gap to whoever leads that scope.
Human preference37.4−30.1#279/3421/1
Human preference37.4−30.1#279/3421/1
Safety37.8−23.7#332/3372/3
Safe-prompt compliance26−34.4#80/821/1
Jailbreak resistance51.6−13.3#165/2721/1
Secure code46.6−19.9#186/2741/1
Toxicity avoidance46.4−12.6#226/2721/1
Fairness40.9−29.1#274/3002/2
Harm refusal32.2−29.4#299/3002/2
§ 2 · Head to head
What it beats, and what beats it
The same models turn up scope after scope. Counted once: where both were placed, who finished higher. Bars run right for DBRX Instruct, left for the other.
§ 3 · Sources
Where the numbers come from
3 publications, 12 figures. Every one links to the page it was read from.
LMArena Text 1196
HELM Safety · HarmBench 27.1%HELM Safety · SimpleSafetyTests 53.5%HELM Safety · Anthropic Red Team 76.6%HELM Safety · BBQ 79.2%HELM Safety · XSTest 77.4%
Enkrypt · Jailbreak risk 11.4%Enkrypt · Harmful content risk 65.6%Enkrypt · CBRN risk 14.8%Enkrypt · Toxicity risk 8.8%Enkrypt · Bias risk 88.1%Enkrypt · Insecure code risk 40.4%
Badge
[](https://publicai.io/model-index/m/dbrx-instruct)