Command R+ vs GPT-4o, Same Price, Same Context, Very Different 2024

Command R+GPT-4o
arenaElo12131345
humanEval71.590.2
sweBench2033
mmlu75.788.7
gpqa4253.6
aime2542
Price in/out ($/M)$2.5 / $10$2.5 / $10

Our verdict

These two are unusually easy to compare because the commercial terms are identical: $2.5 / $10 per M tokens and a 128K context window on both. Everything else favours GPT-4o, which leads on every metric the two publish, among them MMLU 88.7% against 75.7%, HumanEval 90.2% against 71.5%, SWE-bench Verified 33% against 20% and 1345 Arena Elo against 1213, and it adds audio on top of text and vision where Command R+ is text only. The more useful conclusion is the one neither vendor will print: both are 2024 models being sold at 2024 prices, and in 2026 that same $2.5 buys far more elsewhere. If you are picking between these two today, the real answer is that your upgrade path matters more than this duel.

Command R+ · GPT-4o · All AI model comparisons