‹ PublicAI Index
The LLM benchmark aggregator.
Nova 2 Lite
Amazon
Strongest in Factual grounding (#11 of 101), weakest in Agents (#196 of 267). Above par in 4 of 7 scopes. Among the models it meets almost everywhere, it finishes behind GPT-5.6 Terra and Claude Sonnet 5 and ahead of Gemma 3 27B It and Granite 3.1 8B Instruct.
§ 1 · Profile
What it is good at
Bars run from 50 — the average of the models each source lists — so right of the line is above par. The middle column is the gap to whoever leads that scope.
Safety53.8−7.7#74/3361/3
Factual grounding62.3−8.3#11/1011/1
Human preference50.8−16.7#185/3411/1
Human preference50.8−16.7#185/3411/1
Agents44.1−24#196/2672/5
Tool use42.2−31.8#58/811/1
Knowledge work43.8−29.8#111/1771/2
§ 2 · Head to head
What it beats, and what beats it
The same models turn up scope after scope. Counted once: where both were placed, who finished higher. Bars run right for Nova 2 Lite, left for the other.
§ 3 · Sources
Where the numbers come from
4 publications, 4 figures. Every one links to the page it was read from.
LMArena Text 1335
GDPval-AA 354
BFCL v4 27.1%
Vectara · Factual consistency 94.9%
Badge
[](https://publicai.io/model-index/m/nova-2-lite)