DeepSeek V4 vs Claude Opus 4.6: Open-Source Catches the Previous Frontier King

DeepSeek V4Claude Opus 4.6
arenaElo13951402
humanEval9191.4
sweBench6253.4
mmlu91.590.8
gpqa7362.1
aime9084
Price in/out ($/M)$1.74 / $3.48$15 / $75

Our verdict

DeepSeek V4 (April 2026) effectively retires the case for Opus 4.6 as a budget-frontier choice: V4 ties or beats Opus 4.6 on MMLU and GPQA, doubles the context (1M vs 500K) and ships at ~12% of Opus 4.6's API cost ($1.74/$3.48 vs $15/$75 per M tokens). Opus 4.6 still wins on Chatbot Arena, agentic tool reliability and Anthropic's safety tuning. If you can self-host or accept Chinese-vendor cloud, V4 is the rational replacement; for regulated enterprise pipelines already on Anthropic, stick with Opus 4.6 (or upgrade straight to 4.7).

DeepSeek V4 · Claude Opus 4.6 · All AI model comparisons