‹ PublicAI Index
The LLM benchmark aggregator.
Ling 3.0 Flash VL
InclusionAI · 125B · open weights
Strongest in Knowledge work (#46 of 178), weakest in Safety (#283 of 337). Above par in 4 of 10 scopes. Among the models it meets almost everywhere, it finishes behind Mimo V2.6 Pro and Grok 4.7 and ahead of Grok 4 and Gemma 3 4B It.
§ 1 · Profile
What it is good at
Bars run from 50 — the average of the models each source lists — so right of the line is above par. The middle column is the gap to whoever leads that scope.
Agents54.8−13.3#68/2682/5
Knowledge work56.2−17.4#46/1782/2
Core abilities49.7−17.7#104/2041/3
General intelligence49.6−20.5#106/2041/3
Safety45.5−16#283/3371/3
Toxicity avoidance55.1−3.9#106/2721/1
Secure code51.7−14.8#140/2741/1
Harm refusal45.6−16#234/3001/2
Jailbreak resistance32.1−32.8#249/2721/1
Fairness41.9−28.1#264/3001/2
§ 2 · Head to head
What it beats, and what beats it
The same models turn up scope after scope. Counted once: where both were placed, who finished higher. Bars run right for Ling 3.0 Flash VL, left for the other.
§ 3 · Sources
Where the numbers come from
4 publications, 9 figures. Every one links to the page it was read from.
Artificial Analysis Intelligence Index 25
GDPval-AA 1150
AA-Briefcase 984
Enkrypt · Jailbreak risk 27.9%Enkrypt · Harmful content risk 5.6%Enkrypt · CBRN risk 52.2%Enkrypt · Toxicity risk 2.7%Enkrypt · Bias risk 90.4%Enkrypt · Insecure code risk 30.2%
Badge
[](https://publicai.io/model-index/m/ling-3-0-flash-vl)