GLM 5.1
glm-5.1GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on... Built by Z.ai.
Prices updated
Input price
$0.97
per 1M tokens · Standard
Output price
$3.04
per 1M tokens · Standard
Input limit
205K
tokens
Output limit
128K
tokens
Input formats
Output formats
GLM 5.1 price history
6 price records since 9/25/2026 · 4 changes
GLM 5.1 cost calculator
$35/month
Standard pricing
Overview
What is GLM 5.1?
GLM 5.1 is a text & reasoning model from Z.ai. GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on... Its 205K context window and $0.97 input price make it a candidate for quality-focused production applications.
Benchmarks
GLM 5.1 benchmarks & speed
How GLM 5.1 scores on standardized evaluations, and where it lands among every model we track.
26.1
Intelligence Index
58tok/s
Output speed
Reasoning
GPQA Diamond
86.8%
Graduate-level scientific reasoning · top 35%
HLE
30.1%
Humanity's Last Exam · top 36%
Coding
SciCode
44.8%
Python for scientific computing · top 71%
Latency & design
- Time to first token
- 1.77s
Independent scores from Artificial Analysis and DesignArena · updated 10/5/2026. Higher is better; ranks compare against every model we track with that score.
Rates
GLM 5.1 pricing
Pricing is based on token usage. Provider-specific caching, batch, regional, and tool charges may affect the final cost.
Input · Standard
$0.97
Per 1M tokens
Cached input · Standard
$0.18
Per 1M tokens
Capabilities
GLM 5.1 Tools
Tools available when using GLM 5.1 through supported provider APIs.
Function calling
Supported
Structured outputs
Supported
JSON mode
Supported
Reasoning
Supported
Built-in web search
Not supported
Log probabilities
Supported
Deterministic seed
Supported
Parallel tool calls
Not supported
Prompt caching
Supported
Strengths and limitations
Strengths
- Measured results (AA Intelligence 26.1) position it for demanding production work.
- 205K context supports large documents, repositories and extended conversations.
- Supports text & reasoning workloads in one model.
Limitations
- Generated tokens cost 3× more than input tokens, which matters for verbose responses.
- Large context capacity does not guarantee consistent retrieval across the entire prompt.
- Arena Elo and MMLU-Pro are directional; test accuracy, latency and reliability on your own workload before committing.
Arena Elo, MMLU-Pro and pricing figures are illustrative; validate current vendor terms before purchase.
Same provider
