0

Claude 4 Sonnet vs GLM 5.1

Head to head on 20 shared benchmarks, with price, context window, and release dates.

Benchmark scores

55

evals · sort by any column, ⤢ to expand fullscreen

Model
Claude 4 Sonnet

Anthropic

54.7%1480elo58.468.3%4.3%45.4%64.8%77.5%54.6%44.3%72.9%70.5%69.0%33.9%72.4%83.7%37.3%69.6%27.3%52.3%40.7%38.0%---22.9%47.8%-19.8%---67.5%-44.9%93.4%-86.7%62.5%43.9%-6.186.1835.0%--58.3%35.6%74.6%-0.59--84.0%84.0%55.5%
GLM 5.1

Zai

64.5%1589elo73.583.9%27.9%52.0%72.7%75.4%63.2%68.5%71.8%84.9%72.5%41.6%72.3%85.4%36.1%71.2%35.6%97.1%--20.8%35.9%1.13--40.0%-44.8%33.5%34.7%-15.1%--67.0%---22.2%---34.6%5.24---69.8-52.5%31.5%--53.6%
2 / 2 models
AttributeClaude 4 Sonnet

Anthropic

GLM 5.1

Zai

DescriptionClaude 4 Sonnet is an AI model from Anthropic.GLM-5.1 is an AI model from Zai, released with open weights.
FamilyClaudeGlm
Params--
Open weightsclosedopen
Licenseproprietarymit
Released22 May 20257 Apr 2026
Context200,000204,800
Input $/1M$3.00$1.39
Output $/1M$15.00$4.40
Cache read $/1M-$0.234
Cache write $/1M--
Speed0 tok/s0 tok/s
Latency0.00 s0.00 s
Modalitiestext, imagetext
Providers--
PublisherAnthropicZai
Reported scores

FAQ

Is Claude 4 Sonnet or GLM 5.1 better?
Across the 20 benchmarks both models report on Sophon, GLM 5.1 leads on 16 of them. Which one is "better" depends on the benchmark - the table above shows every shared score.
How do Claude 4 Sonnet and GLM 5.1 compare on GPQA Diamond?
Claude 4 Sonnet scores 68.3% and GLM 5.1 scores 83.9% on GPQA Diamond.
Which is cheaper, Claude 4 Sonnet or GLM 5.1?
GLM 5.1 is cheaper at $4.40 per million output tokens against $15.00 for Claude 4 Sonnet - about 3.4x.
When were Claude 4 Sonnet and GLM 5.1 released?
Claude 4 Sonnet was released 22 May 2025 by Anthropic. GLM 5.1 was released 7 Apr 2026 by Zai.

Comparing something else? Build your own side-by-side.