‹ PublicAI Index
The LLM benchmark aggregator.
Ox Alpha
Other
Strongest in Data analysis (#28 of 57), weakest in Core abilities (#188 of 204). Above par in 1 of 11 scopes. Among the models it meets almost everywhere, it finishes behind Claude Fable 5 and GPT-6 Astra and ahead of MiniMax M3 and Qwen3.6 Plus.
§ 1 · Profile
What it is good at
Bars run from 50 — the average of the models each source lists — so right of the line is above par. The middle column is the gap to whoever leads that scope.
Coding48.7−20.4#84/1651/5
Code generation46.7−23.3#54/771/2
Agentic coding50−18.7#74/1571/4
Reasoning40.8−26.1#150/1782/4
Reasoning45.2−23.9#81/1402/3
Mathematics32.9−32.8#138/1401/2
Core abilities39.5−27.9#188/2041/3
Data analysis54−12.1#28/571/1
Instruction following36.6−37.1#48/571/1
Language26−45.5#55/571/1
General intelligence41−29.1#165/2041/3
§ 2 · Head to head
What it beats, and what beats it
The same models turn up scope after scope. Counted once: where both were placed, who finished higher. Bars run right for Ox Alpha, left for the other.
§ 3 · Sources
Where the numbers come from
2 publications, 9 figures. Every one links to the page it was read from.
LiveBench 69.2LiveBench · Reasoning 76.6LiveBench · Coding 75.8LiveBench · Agentic Coding 52.6LiveBench · Mathematics 77.5LiveBench · Data Analysis 75.8LiveBench · Language 66.1LiveBench · Instruction Following 60.3
SimpleBench 55.1%
Badge
[](https://publicai.io/model-index/m/ox-alpha)