Amazon model benchmarks.

Compare Amazon models across answer quality, coding, complex tasks, speed, price, and context size.

Amazon

Quality

Highest overall scores

  1. 1Amazon: Nova 2 Lite18.2
  2. 2Amazon: Nova Lite 1.00
  3. 3Amazon: Nova Premier 1.00

Response speed

Fastest output in this sample

  1. 1Amazon: Nova 2 Lite102 t/s
  2. 2Amazon: Nova Lite 1.080 t/s
  3. 3Amazon: Nova Premier 1.017.5 t/s

Input price

Lowest price per 1M tokens

  1. 1Amazon: Nova Lite 1.0$0.06/1M
  2. 2Amazon: Nova 2 Lite$0.30/1M
  3. 3Amazon: Nova Premier 1.0$2.50/1M

Highlights

Quality

Overall capability · Higher is better

Speed

Output tokens per second · Higher is better

Input price

USD per 1M tokens · Lower is better

Quality benchmarks

Compare Amazon model capability.

Performance and price

Compare practical tradeoffs.

All Amazon models.

3 models in the current catalogue.

Amazon technical benchmarks.

Compare named evaluation records available for Amazon models. Missing results remain blank rather than being estimated.

BenchmarkAreaNova 2 LiteNova Lite 1.0Nova Premier 1.0Source
Intelligence IndexoverallQuality18.2OpenRouter / artificial-analysis
Coding Indexcoding23OpenRouter / artificial-analysis
Agentic Indexagentic3.1OpenRouter / artificial-analysis
GPQAGPQA Diamond81.143.356.9OpenRouter model benchmarks
Humanity's Last ExamHLE10.94.64.7OpenRouter model benchmarks
IFBenchIFBench70.734.136.2OpenRouter model benchmarks
τ²-Bench Telecomτ²-Bench Telecom72.817.538.3OpenRouter model benchmarks
AA-LCRAA-LCR55.317.730OpenRouter model benchmarks
GDPval-AAGDPval-AA4.6OpenRouter model benchmarks
CritPtCritPt0.300OpenRouter model benchmarks
SciCodeSciCode36.913.927.9OpenRouter model benchmarks
Terminal-Bench HardTerminal-Bench Hard16.70.86.8OpenRouter model benchmarks
AA-Omniscience AccuracyAA-Omniscience Accuracy18.89.719.1OpenRouter model benchmarks
AA-Omniscience Non-Hallucination RateAA-Omniscience Non-Hallucination Rate10.244.431.5OpenRouter model benchmarks