Claude Sonnet 4.5 vs GPT-5 Mini
Head to head on 31 shared benchmarks, with price, context window, and release dates.
Benchmark scores
116evals · sort by any column, ⤢ to expand fullscreen
| Model | |||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| GPT-5 Mini OpenAI | 46.7% | 60.2% | 1310elo | 56.2 | 27.2% | 68.7% | 5.1% | 45.6% | 69.1% | 68.2% | 55.2% | 65.3% | 75.5% | 82.2% | 68.3% | 54.5% | 1.54 | 1.54 | 43.0% | 80.6% | 77.5% | 66.9% | 9.0% | 43.0% | 36.9% | 39.7% | 59.8% | 75.2% | 14.4% | 14.2% | 31.9% | - | - | 0.0% | - | - | - | 0.0% | - | 1.13 | - | 0.0% | -0.15 | 20.0% | 25.0% | - | 40.0% | 1.7 | - | - | 100.0% | 18.1% | 2.8% | 100.0% | 66.7% | - | - | - | - | 0.0% | 88.3% | - | 60.0% | - | - | 100.0% | - | - | - | 90.0% | - | - | 47.2% | - | 80.0% | - | - | - | 0.0% | - | - | - | 80.0% | - | - | - | - | - | - | - | - | - | - | - | 16.8% | 4.95 | - | - | - | - | - | - | - | - | 0.0% | 9.5% | 0.0% | - | 86.1% | 0.49 | 39.21 | 272.11 | - | 86.0% | 86.0% | - | 47.9% |
| Claude Sonnet 4.5 Anthropic | 37.0% | 60.8% | 1675elo | 71.4 | 15.2% | 72.7% | 7.2% | 42.7% | 70.7% | 80.4% | 57.0% | 53.4% | 76.5% | 79.3% | 77.6% | 59.0% | 0.75 | 0.75 | 40.6% | 84.5% | 86.0% | 64.0% | 19.0% | 32.9% | 42.8% | 67.0% | 74.8% | 73.3% | 28.8% | 22.6% | 70.5% | 0.0% | 0.0% | - | 100 | 0.0% | 0.0% | - | 0.0% | - | 5.0% | - | - | - | - | 0.0% | - | - | 0.0% | 0.0% | - | - | - | - | - | 0.0% | 0.0% | 0.0% | 0.0% | - | - | 0.0% | - | 0.0% | 83.9% | - | 0.0% | 0.0% | 0.0% | - | 0.0% | 0.0% | - | 87.5% | - | 83.8% | 0.0% | 0.0% | - | 0.0% | 0 | 0 | - | 0.0% | 0.0% | 62.9% | 0.0% | 0.0% | 0.0% | 0.0% | 0.0% | 0.0% | 0.0% | 0.0% | - | - | 0 | 0.0% | 0.0% | 77.2% | 0.0% | 0 | 0 | 50.34 | - | - | - | 0 | - | - | - | - | 0.0% | - | - | 1.49 | 27.5% |
2 / 2 models
| Attribute | Claude Sonnet 4.5 Anthropic | GPT-5 Mini OpenAI |
|---|---|---|
| Description | anthropic/claude-sonnet-4.5 is an AI model. | GPT-5 mini is an AI model from OpenAI. |
| Family | Claude | GPT |
| Params | Undisclosed (closed) | - |
| Open weights | closed | closed |
| License | Proprietary | proprietary |
| Released | 29 Sep 2025 | 7 Aug 2025 |
| Context | 1,000,000 | 400,000 |
| Input $/1M | $3.00 | $0.25 |
| Output $/1M | $15.00 | $2.00 |
| Cache read $/1M | $0.300 | $0.025 |
| Cache write $/1M | $3.750 | - |
| Speed | 0 tok/s | 0 tok/s |
| Latency | 0.00 s | 0.00 s |
| Modalities | text, image, file | text, image, file |
| Providers | anthropic, aws-bedrock, vertex-ai | - |
| Publisher | Anthropic | OpenAI |
| Reported scores | TaxEval v273.3% LiveBench - Math79.3% (LiveBench) LiveBench - Coding80.4% (LiveBench) LiveBench - Language76.5% (LiveBench) LiveBench - Data Analysis57.0% (LiveBench) |
FAQ
- Is Claude Sonnet 4.5 or GPT-5 Mini better?
- Across the 31 benchmarks both models report on Sophon, Claude Sonnet 4.5 leads on 20 of them. Which one is "better" depends on the benchmark - the table above shows every shared score.
- How do Claude Sonnet 4.5 and GPT-5 Mini compare on GPQA Diamond?
- Claude Sonnet 4.5 scores 72.7% and GPT-5 Mini scores 68.7% on GPQA Diamond.
- Which is cheaper, Claude Sonnet 4.5 or GPT-5 Mini?
- GPT-5 Mini is cheaper at $2.00 per million output tokens against $15.00 for Claude Sonnet 4.5 - about 7.5x.
- When were Claude Sonnet 4.5 and GPT-5 Mini released?
- Claude Sonnet 4.5 was released 29 Sep 2025 by Anthropic. GPT-5 Mini was released 7 Aug 2025 by OpenAI.
Comparing something else? Build your own side-by-side.