GLM-5 vs GPT-5.1 Codex

Side-by-side benchmark comparison across coding, math, reasoning, speed, and pricing.

GLM-5 by Zhipu AI wins on 7 of 10 benchmarks against GPT-5.1 Codex by OpenAI, which leads on 3. This head-to-head comparison covers coding, math, reasoning, speed, and pricing metrics from our benchmark data.

Category-by-Category Breakdown

General Intelligence: In general intelligence, GLM-5 scores 1456 on Chatbot Arena ELO compared to GPT-5.1 Codex's 1395.

Coding: In coding, GLM-5 scores 1235 on Design Arena ELO compared to GPT-5.1 Codex's 1197, while GLM-5 scores 1260 on Website Arena ELO compared to GPT-5.1 Codex's 1174.

Reasoning: In reasoning, GLM-5 scores 78.6% on GPQA Diamond compared to GPT-5.1 Codex's 65.0%.

Context: In context, GPT-5.1 Codex scores 400K on Context Length compared to GLM-5's 205K.

Pricing Comparison

GLM-5 costs $0.60/1M input tokens and $1.9/1M output tokens, while GPT-5.1 Codex costs $1.3/1M input and $10.0/1M output. GLM-5 is the more affordable option for API usage.

Speed Comparison

GLM-5 generates output at 55 tok/s compared to GPT-5.1 Codex's 85 tok/s, and the time to first token is 1030 ms for GLM-5 versus 400 ms for GPT-5.1 Codex. GPT-5.1 Codex delivers faster throughput.

Verdict

For developers prioritizing general intelligence and affordability, GLM-5 has the edge. For those who value speed, GPT-5.1 Codex is the stronger choice.

View Individual Model Pages

GLM-5 vs GPT-5.1 Codex — FAQ

Which is better, GLM-5 or GPT-5.1 Codex?

GLM-5 wins on more benchmarks overall (7 vs 3). However, the best choice depends on your specific needs — each model excels in different areas.

How does GLM-5 compare to GPT-5.1 Codex for coding?

SWE-bench Verified data is not available for both models. Check the detailed comparison charts above for other coding-related metrics.

Is GLM-5 cheaper than GPT-5.1 Codex?

Yes, GLM-5 is cheaper. GLM-5 costs $0.60/1M input and $1.9/1M output tokens. GPT-5.1 Codex costs $1.3/1M input and $10.0/1M output tokens.

Which is faster, GLM-5 or GPT-5.1 Codex?

GPT-5.1 Codex is faster, generating output at 85 tok/s compared to 55 tok/s. Faster output speed means shorter wait times for API responses.

What benchmarks does the GLM-5 vs GPT-5.1 Codex comparison cover?

This comparison covers 10 benchmarks including Chatbot Arena ELO, GPQA Diamond, Output Speed, Time to First Token, Input Cost, Output Cost, Context Length, Cached Input Cost, and more. Metrics span general intelligence, coding, math, reasoning, speed, and cost categories.