0

Claude Opus 4.6 vs GPT-4.1

Head to head on 13 shared benchmarks, with price, context window, and release dates.

Benchmark scores

64

evals · sort by any column, ⤢ to expand fullscreen

Model
Claude Opus 4.6

Anthropic

2.0867.0%1804elo40.7%84.0%19.1%44.6%68.5%45.7%75.6%76.0%48.5%84.8%----99.79--56.1%1.5--71.0%-1223elo77.7-84.9%887715878.8%78.2%69.9%63.3%83.3%89.3%88.7%-----78.5%--48.2%-86.7%--9350.0%51.6%-44.572.0%--80.657.6%--66.5%
GPT-4.1

OpenAI

1.5563.1%1417elo5.5%66.6%4.2%43.0%65.9%38.1%39.6%75.1%13.6%47.1%0.0%52.4%43.7%34.7%-26.7%61.5%--1.6483.7%-12.4%--60.9%----------45.7%70.8%0.490.4991.3%-32.1%32.1%-69.4%-80.6%50.5%---8.57--31.1%10.0%--92.0%92.0%48.0%
2 / 2 models
AttributeClaude Opus 4.6

Anthropic

GPT-4.1

OpenAI

DescriptionClaude Opus 4.6 is an AI model from Anthropic.GPT-4.1 is an AI model from OpenAI.
FamilyClaudeGPT
Params--
Open weightsclosedclosed
Licenseproprietaryproprietary
Released5 Feb 202614 Apr 2025
Context1,000,0001,047,576
Input $/1M$5.00$2.00
Output $/1M$25.00$8.00
Cache read $/1M$0.500$0.500
Cache write $/1M$6.250-
Speed0 tok/s0 tok/s
Latency0.00 s0.00 s
Modalitiestext, image, fileimage, text, file
Providers--
PublisherAnthropicOpenAI
Reported scores

FAQ

Is Claude Opus 4.6 or GPT-4.1 better?
Across the 13 benchmarks both models report on Sophon, Claude Opus 4.6 leads on 13 of them. Which one is "better" depends on the benchmark - the table above shows every shared score.
How do Claude Opus 4.6 and GPT-4.1 compare on GPQA Diamond?
Claude Opus 4.6 scores 84.0% and GPT-4.1 scores 66.6% on GPQA Diamond.
Which is cheaper, Claude Opus 4.6 or GPT-4.1?
GPT-4.1 is cheaper at $8.00 per million output tokens against $25.00 for Claude Opus 4.6 - about 3.1x.
When were Claude Opus 4.6 and GPT-4.1 released?
Claude Opus 4.6 was released 5 Feb 2026 by Anthropic. GPT-4.1 was released 14 Apr 2025 by OpenAI.

Comparing something else? Build your own side-by-side.