0

Claude 4 Sonnet vs GPT-4o-mini

Head to head on 18 shared benchmarks, with price, context window, and release dates.

Benchmark scores

58

evals · sort by any column, ⤢ to expand fullscreen

Model
Claude 4 Sonnet

Anthropic

40.7%38.0%54.7%1480elo68.3%4.3%45.4%64.8%77.5%44.3%72.9%44.9%93.4%83.7%62.5%37.3%69.6%0.59---22.9%47.8%--19.8%-58.467.5%----54.6%70.5%69.0%33.9%-72.4%86.7%-43.9%6.186.1835.0%-58.3%35.6%74.6%-27.3%---84.0%84.0%-52.3%55.5%
GPT-4o-mini

OpenAI

11.7%14.7%45.5%871elo42.6%4.2%31.0%47.2%43.0%65.8%32.9%23.4%78.9%64.8%54.5%22.9%60.5%0.3616.7%3.6%50.7%--0.290.0%-0.0%--0.0%1.540.0%72.0%----62.8%--78.0%----3.81---100.0%-0.7883.3%54.5%--22.5%-39.6%
2 / 2 models
AttributeClaude 4 Sonnet

Anthropic

GPT-4o-mini

OpenAI

DescriptionClaude 4 Sonnet is an AI model from Anthropic.GPT-4o Mini is an AI model from OpenAI.
FamilyClaudeGPT
Params--
Open weightsclosedclosed
Licenseproprietaryproprietary
Released22 May 202518 Jul 2024
Context200,000128,000
Input $/1M$3.00$0.15
Output $/1M$15.00$0.60
Cache read $/1M-$0.075
Cache write $/1M--
Speed0 tok/s0 tok/s
Latency0.00 s0.00 s
Modalitiestext, imagetext, image, file
Providers--
PublisherAnthropicOpenAI
Reported scores

FAQ

Is Claude 4 Sonnet or GPT-4o-mini better?
Across the 18 benchmarks both models report on Sophon, Claude 4 Sonnet leads on 17 of them. Which one is "better" depends on the benchmark - the table above shows every shared score.
How do Claude 4 Sonnet and GPT-4o-mini compare on GPQA Diamond?
Claude 4 Sonnet scores 68.3% and GPT-4o-mini scores 42.6% on GPQA Diamond.
Which is cheaper, Claude 4 Sonnet or GPT-4o-mini?
GPT-4o-mini is cheaper at $0.60 per million output tokens against $15.00 for Claude 4 Sonnet - about 25.0x.
When were Claude 4 Sonnet and GPT-4o-mini released?
Claude 4 Sonnet was released 22 May 2025 by Anthropic. GPT-4o-mini was released 18 Jul 2024 by OpenAI.

Comparing something else? Build your own side-by-side.