0

Claude Opus 4.5 vs GPT-4.1

Head to head on 16 shared benchmarks, with price, context window, and release dates.

Benchmark scores

63

evals · sort by any column, ⤢ to expand fullscreen

Model
Claude Opus 4.5

Anthropic

62.7%61.3%1683elo20.7%84.7%81.0%13.2%43.0%73.8%88.9%68.7%47.0%79.2%74.3%40.9%86.3%------16.7%--35.6%6.1%-73.112.1383.8%83.5%78.1%79.7%74.4%62.5%81.3%90.4%80.1%------45.2%-83.2%-100.0%37.436.0%77.5%45.0%-45.370.7%--61.820.6%--62.2%
GPT-4.1

OpenAI

34.7%63.1%1417elo5.5%60.9%66.6%4.2%43.0%45.7%80.6%65.9%38.1%39.6%75.1%13.6%47.1%0.0%52.4%43.7%26.7%61.5%1.64-83.7%1.55--12.4%-----------70.8%0.490.4991.3%32.1%32.1%-69.4%-50.5%-----8.57--31.1%10.0%--92.0%92.0%48.0%
2 / 2 models
AttributeClaude Opus 4.5

Anthropic

GPT-4.1

OpenAI

DescriptionClaude Opus 4.5 is an AI model from Anthropic.GPT-4.1 is an AI model from OpenAI.
FamilyClaudeGPT
Params--
Open weightsclosedclosed
Licenseproprietaryproprietary
Released24 Nov 202514 Apr 2025
Context200,0001,047,576
Input $/1M$5.00$2.00
Output $/1M$25.00$8.00
Cache read $/1M$0.500$0.500
Cache write $/1M$6.250-
Speed0 tok/s0 tok/s
Latency0.00 s0.00 s
Modalitiesfile, image, textimage, text, file
Providers--
PublisherAnthropicOpenAI
Reported scores

FAQ

Is Claude Opus 4.5 or GPT-4.1 better?
Across the 16 benchmarks both models report on Sophon, Claude Opus 4.5 leads on 13 of them. Which one is "better" depends on the benchmark - the table above shows every shared score.
How do Claude Opus 4.5 and GPT-4.1 compare on GPQA Diamond?
Claude Opus 4.5 scores 81.0% and GPT-4.1 scores 66.6% on GPQA Diamond.
Which is cheaper, Claude Opus 4.5 or GPT-4.1?
GPT-4.1 is cheaper at $8.00 per million output tokens against $25.00 for Claude Opus 4.5 - about 3.1x.
When were Claude Opus 4.5 and GPT-4.1 released?
Claude Opus 4.5 was released 24 Nov 2025 by Anthropic. GPT-4.1 was released 14 Apr 2025 by OpenAI.

Comparing something else? Build your own side-by-side.