Introducing GLM-5Z.AIReleased February 12, 2026

GLM-5

Fifth-generation foundation model shifting from coding to complex systems engineering, benchmarked against Claude Opus 4.5.

reasoning-llmProprietary API$1.00 / 1M tok (DeepSeek Sparse Attention)Context: 200K (200,000 tokens)Arena ELO: 1420

Technical Specifications

Architecture Type
DeepSeek Sparse Attention Transformer
Total Parameters
Fifth-Gen Scale
Context Window
200K (200,000 tokens)
Max Output Tokens
64K (64,000 tokens)
Knowledge Cutoff
January 2026
Supported Modalities
text, code, reasoning
License & Access
Z.AI Terms of Service

Benchmark Evaluations

Mmlu Pro
84.5
Humaneval
94.2
Swe Bench Lite
62
Chatbot Arena ELO
1420

Deep Architectural Overview

GLM-5 integrates DeepSeek Sparse Attention for superior token efficiency and KV cache preservation. It demonstrates deep reasoning performance in backend architecture, complex algorithmic optimization, and stubborn bug fixing over 200k context.

Strengths & Considerations

Core Strengths
  • DeepSeek Sparse Attention integration for token efficiency
  • Directly aligned with Claude Opus 4.5 in code-logic density
  • Exceptional backend architecture and debugging capabilities
Known Limitations
  • Succeeded by GLM-5.2 for 1M context workflows

Token & API Pricing

Input Tokens (1M)$1.00
Output Tokens (1M)$3.20
Cached Input (1M)$0.2000
Pricing is verified directly against Z.AI's developer documentation and API rate sheets.

Similar & Alternative Models

Explore other frontier models from Z.AI and comparable reasoning engines.

Browse all models
Z.AI

Ultra-low latency version of GLM-5.3-Flash engineered for high-concurrency real-time IDE code completion and interactive agents.

200K ctx$0.37 / 1M tok (Ultra-Low Latency MoE)
Z.AI

The first native multimodal model in the GLM-5 series: 320B parameters (18B active) combining linear and sparse attention for low-cost visual coding.

200K ctx$0.15 / 1M tok (320B MoE Visual Coding)
Z.AI

Z.AI’s premier flagship model, delivering a 50% performance gain on Code Bench and matching Claude Mythos 5 in cybersecurity and vulnerability discovery.

1M ctx$1.40 / 1M tok (Flagship Coding & Cyber)
Z.AI

Flagship model built for project-scale context, supporting truly usable 1M-token context with 128k output.

1M ctx$1.40 / 1M tok (1M Lossless Context)