glm-5.1

Common Name: GLM-5.1

ChatGLM
Released on Apr 8 12:00 AMKnowledge Cutoff Apr 1, 2025 12:00 AMSupportedTool InvocationSupportedReasoning
CompareTry in Chat

GLM-5.1 is Zhipu AI's flagship model aligned with Claude Opus 4.6 in capability, featuring an 8-hour long-horizon agent runtime and top-tier performance on complex engineering tasks.

Specifications

Context
200K
Maximum Output
128K
Inputtext
Outputtext

Performance (7-day Average)

Collecting…
Collecting…
Collecting…

Pricing

< 32K
Input¥6.60/MTokens
Output¥26.40/MTokens
Cached Input¥1.43/MTokens
32K-200K
Input¥8.80/MTokens
Output¥30.80/MTokens
Cached Input¥2.20/MTokens

Performance Metrics (24h)

Similar Models

¥4.40/¥13.20/M
ctx128Kmaxavailtps
InOutCap

Zhipu AI's GLM-4.5 AirX variant optimized for high-speed inference.

¥8.80/¥17.60/M
ctx128Kmaxavailtps
InOutCap

Zhipu AI's GLM-4.5 X variant with enhanced performance.

¥4.40/¥19.80/M
ctx200Kmax128Kavailtps
InOutCap

GLM-5 is Zhipu AI's new-generation flagship base model (744B total / 40B active MoE), with significantly improved coding, reasoning, and agentic capabilities over GLM-4.7.

¥5.50/¥24.20/M
ctx200Kmax128Kavailtps
InOutCap

GLM-5-Turbo is the fast variant of GLM-5, optimized for lower latency and cost while retaining strong agentic and coding performance.