Anthropic model benchmarks.

Compare Anthropic models across answer quality, coding, complex tasks, speed, price, and context size.

Quality

Highest overall scores

  1. 1Claude Opus 560.7
  2. 2Anthropic: Claude Sonnet 553.4
  3. 3Claude Opus 5 (Fast)0

Response speed

Fastest output in this sample

  1. 1Claude Opus 5 (Fast)112 t/s
  2. 2Anthropic: Claude Sonnet 596.5 t/s
  3. 3Claude Opus 563 t/s

Input price

Lowest price per 1M tokens

  1. 1Anthropic: Claude Sonnet 5$2.00/1M
  2. 2Claude Opus 5$5.00/1M
  3. 3Claude Opus 5 (Fast)$10.00/1M

Highlights

Quality

Overall capability · Higher is better

Speed

Output tokens per second · Higher is better

Input price

USD per 1M tokens · Lower is better

Quality benchmarks

Compare Anthropic model capability.

Performance and price

Compare practical tradeoffs.

All Anthropic models.

3 models in the current catalogue.

ModelLinks
Claude Opus 560.763 t/s1M$5.00$25.00
Anthropic: Claude Sonnet 553.496.5 t/s1M$2.00$10.00
Claude Opus 5 (Fast)0112 t/s1M$10.00$50.00

Anthropic technical benchmarks.

Compare named evaluation records available for Anthropic models. Missing results remain blank rather than being estimated.

BenchmarkAreaClaude Opus 5Claude Sonnet 5Claude Opus 5 (Fast)Source
Intelligence IndexoverallQuality60.753.4OpenRouter / artificial-analysis
Coding Indexcoding7871.5OpenRouter / artificial-analysis
Agentic Indexagentic55.346.7OpenRouter / artificial-analysis
GPQAGPQA Diamond93.291.1OpenRouter model benchmarks
Humanity's Last ExamHLE52.639.6OpenRouter model benchmarks
AA-LCRAA-LCR7070.7OpenRouter model benchmarks
GDPval-AAGDPval-AA68.155.2OpenRouter model benchmarks
CritPtCritPt29.116.9OpenRouter model benchmarks
SciCodeSciCode55.753.6OpenRouter model benchmarks
AA-Omniscience AccuracyAA-Omniscience Accuracy54.238.3OpenRouter model benchmarks
AA-Omniscience Non-Hallucination RateAA-Omniscience Non-Hallucination Rate49.962.7OpenRouter model benchmarks