‹ PublicAI Index
The LLM benchmark aggregator.
Cogito 671B V2 P1
DeepSeek · 671B
Strongest in Secure code (#18 of 274), weakest in Jailbreak resistance (#135 of 272). Above par in 6 of 6 scopes. Among the models it meets almost everywhere, it finishes behind Apriel 1.5 15B Thinker and Claude Sonnet 5 and ahead of Qwen3.5 397B A17B and GPT-4.1.
§ 1 · Profile
What it is good at
Bars run from 50 — the average of the models each source lists — so right of the line is above par. The middle column is the gap to whoever leads that scope.
Safety55.9−5.6#37/3361/3
Secure code64.1−2.4#18/2741/1
Toxicity avoidance58−1#34/2721/1
Harm refusal55.6−6#61/2991/2
Fairness51.5−18.5#96/2991/2
Jailbreak resistance53.8−11.1#135/2721/1
§ 2 · Head to head
What it beats, and what beats it
The same models turn up scope after scope. Counted once: where both were placed, who finished higher. Bars run right for Cogito 671B V2 P1, left for the other.
§ 3 · Sources
Where the numbers come from
1 publication, 6 figures. Every one links to the page it was read from.
Enkrypt · Jailbreak risk 9.6%Enkrypt · Harmful content risk 11.7%Enkrypt · CBRN risk 11.7%Enkrypt · Toxicity risk 0.7%Enkrypt · Bias risk 75.5%Enkrypt · Insecure code risk 4.9%
Badge
[](https://publicai.io/model-index/m/cogito-671b-v2-p1)