Anthropic

3 models · average score 41.6 · US

Official site →

#2Claude Opus 4.8

56.0

Claude Opus 4.8 is Anthropic's reasoning-capable model: $5/$25 per 1M input/output tokens, 1M token context window, 69.2% on SWE-Bench Pro.

#6Claude Sonnet 4.6

36.0

Claude Sonnet 4.6 is Anthropic's reasoning-capable model: $3/$15 per 1M input/output tokens, 1M token context window, 14.9% on SWE-Bench Pro.

#9Claude Haiku 4.5 (latest)

32.8

Claude Haiku 4.5 (latest) is Anthropic's reasoning-capable model: $1/$5 per 1M input/output tokens, 200K token context window, 39.45% on SWE-Bench Pro.

model categories covered

Other providers