‹ PublicAI Index
The LLM benchmark aggregator.
Qwen3.6 Max
Alibaba
Strongest in Finance (#31 of 151), weakest in Legal (#136 of 151). Above par in 5 of 9 scopes. Among the models it meets almost everywhere, it finishes behind Claude Opus 5.5 and Claude Opus 4.6 and ahead of Qwen3.7 Max and Kimi K2.6.
§ 1 · Profile
What it is good at
Bars run from 50 — the average of the models each source lists — so right of the line is above par. The middle column is the gap to whoever leads that scope.
Human preference62.8−4.7#48/3421/1
Human preference62.8−4.7#48/3421/1
Reasoning53.6−13.3#55/1781/4
Reasoning55.7−13.4#38/1401/3
Coding49.7−19.4#78/1651/5
Agentic coding49.6−19.1#75/1571/4
Professional48.1−13.5#105/1681/1
Finance56.3−6.6#31/1511/1
Legal38.4−28.7#136/1511/1
§ 2 · Head to head
What it beats, and what beats it
The same models turn up scope after scope. Counted once: where both were placed, who finished higher. Bars run right for Qwen3.6 Max, left for the other.
§ 3 · Sources
Where the numbers come from
3 publications, 5 figures. Every one links to the page it was read from.
LMArena Text 1460
SimpleBench 63%
Vals · CaseLaw 47.91%Vals · CorpFin 66.47%Vals · SWE-bench Verified 72.8% (Mini-SWE-agent)
Badge
[](https://publicai.io/model-index/m/qwen3-6-max)