0

GLM 4.7 vs gpt-oss-120b

Head to head on 20 shared benchmarks, with price, context window, and release dates.

Benchmark scores

41

evals · sort by any column, ⤢ to expand fullscreen

Model
GLM 4.7

Zai

48.0%46.4%1411elo6666.4%6.4%54.6%60.8%73.1%55.2%35.7%65.2%76.0%59.7%56.2%79.4%35.4%68.8%30.3%94.2%----46.7%28.6%2.4%90.0%-33.3%--32.8%33.3%68.6%-100.0%6.0%---51.9%
gpt-oss-120b

OpenAI

66.7%58.2%959elo35.867.2%5.9%58.3%51.0%60.2%38.8%50.3%48.6%68.9%39.2%70.7%77.5%36.0%71.6%5.3%45.0%41.8%38.9%7.6%13.0%----57.6-1.621.53---80.0%--97.9%26.0%65.6%49.6%
2 / 2 models
AttributeGLM 4.7

Zai

gpt-oss-120b

OpenAI

DescriptionGLM-4.7 is an AI model from Zai, released with open weights.GPT-Oss 120b is an AI model from OpenAI, released with open weights.
FamilyGlmGPT
Params--
Open weightsopenopen
Licensemitapache-2.0
Released22 Dec 20255 Aug 2025
Context204,800131,072
Input $/1M$0.60$0.15
Output $/1M$2.20$0.60
Cache read $/1M$0.080$0.030
Cache write $/1M--
Speed0 tok/s186 tok/s
Latency0.00 s0.50 s
Modalitiestexttext
Providers--
PublisherZaiOpenAI
Reported scores
GPQA Diamond67.2% (AA)
LiveCodeBench70.7% (AA)
MMLU-Pro77.5% (AA)
SciCode36.0% (AA)

FAQ

Is GLM 4.7 or gpt-oss-120b better?
Across the 20 benchmarks both models report on Sophon, GLM 4.7 leads on 12 of them. Which one is "better" depends on the benchmark - the table above shows every shared score.
How do GLM 4.7 and gpt-oss-120b compare on GPQA Diamond?
GLM 4.7 scores 66.4% and gpt-oss-120b scores 67.2% on GPQA Diamond.
Which is cheaper, GLM 4.7 or gpt-oss-120b?
gpt-oss-120b is cheaper at $0.60 per million output tokens against $2.20 for GLM 4.7 - about 3.7x.
When were GLM 4.7 and gpt-oss-120b released?
GLM 4.7 was released 22 Dec 2025 by Zai. gpt-oss-120b was released 5 Aug 2025 by OpenAI.

Comparing something else? Build your own side-by-side.