Zhipu AI

GLM-4.5-Flash — Benchmark Scores, Pricing & Performance Analysis

OPEN SOURCEZhipu AI
Input Cost
$0.00/1M
Output Cost
$0.00/1M
Context Window
131K
Max Output
98K
Knowledge Cutoff
Apr 2025
Accepts
text

GLM-4.5-Flash by Zhipu AI demonstrates competitive pricing. View detailed benchmark data including scores across coding, math, reasoning, speed, and cost metrics.

Specifications last checked against the provider on . Benchmark scores carry their own source, linked beside each number.

Cost Benchmarks

Input Cost
$0.001st of 413

Cost per 1M input tokens

OpenRouter API
Output Cost
$0.001st of 413

Cost per 1M output tokens

OpenRouter API

Context Benchmarks

Context Length
131K212th of 451

Maximum context window size

OpenRouter API

Available from 4 providers

The same model costs different amounts depending on who serves it. You can bring your own key for 1 of these — connect it here.

ProviderInput / 1MOutput / 1MCached inContextYour key
EmpirioLabs AI$0$0200K
UnoRouter$0$0131K
Z.AI$0$0$0131K
Zhipu AI$0$0$0131KSupported

GLM-4.5-Flash — Frequently Asked Questions

How much does GLM-4.5-Flash cost?

GLM-4.5-Flash costs $0.00 per 1M input tokens and $0.00 per 1M output tokens. This makes it one of the more affordable models.

What is the context window of GLM-4.5-Flash?

GLM-4.5-Flash has a context window of 131K tokens. This determines how much text, conversation history, and code the model can process in a single request.

Who created GLM-4.5-Flash?

GLM-4.5-Flash was created by Zhipu AI. It is classified as a open source model in our catalogue.

Is GLM-4.5-Flash open source?

Yes, GLM-4.5-Flash is an open-source model. The model weights are publicly available for download and self-hosting.