‹ PublicAI IndexGoogle
The LLM benchmark aggregator.
Gemini 2.0 Flash Thinking
Strongest in Professional (#78 of 168), weakest in Reasoning (#131 of 178). Above par in 2 of 6 scopes. Among the models it meets almost everywhere, it finishes behind Claude Opus 5.5 and Claude Fable 5.1 and ahead of Claude 3.5 Haiku and Qwen3 Max.
§ 1 · Profile
What it is good at
Bars run from 50 — the average of the models each source lists — so right of the line is above par. The middle column is the gap to whoever leads that scope.
Professional50.2−11.4#78/1681/1
Finance50.5−12.4#86/1511/1
Coding44.4−24.7#118/1641/5
Agentic coding43.2−25.5#118/1571/4
Reasoning44.6−22.3#131/1781/4
Reasoning41.3−27.8#109/1401/3
§ 2 · Head to head
What it beats, and what beats it
The same models turn up scope after scope. Counted once: where both were placed, who finished higher. Bars run right for Gemini 2.0 Flash Thinking, left for the other.
§ 3 · Sources
Where the numbers come from
3 publications, 3 figures. Every one links to the page it was read from.
Aider polyglot 18.2%
SimpleBench 30.7%
Vals · TaxEval 69.79%
Badge
[](https://publicai.io/model-index/m/gemini-2-0-flash-thinking)