Claude Opus 4.8 vs Claude Opus 4.7, same Price, Stronger Agent

Claude Opus 4.8Claude Opus 4.7
arenaElo14351420
humanEval94.593.1
sweBench6764.3
mmlu9392.3
gpqa67.565.2
aime9088
Price in/out ($/M)$5 / $25$5 / $25

Our verdict

Opus 4.8 is a clean upgrade at the same list price, agentic coding rises from 64.3% to 69.2% on SWE-Bench Pro, it leads computer use at 84% on Online-Mind2Web, runs autonomously for longer and reports progress more honestly. Fast mode is now 3× cheaper and the minimum cacheable prompt drops to 1,024 tokens. There is no reason to start a new project on 4.7, go straight to 4.8.

Claude Opus 4.8 · Claude Opus 4.7 · All AI model comparisons