Introducing GLM-5.1Z.AIReleased April 7, 2026

GLM-5.1

Engineered for long-horizon tasks, able to work independently for up to 8 hours in a single run, aligned with Claude Opus 4.6.

reasoning-llmProprietary API$1.40 / 1M tok (8h Autonomous Agent)Context: 200K (200,000 tokens)Arena ELO: 1465

Technical Specifications

Architecture Type
Long-Horizon Autonomous Engineering Transformer
Total Parameters
Fifth-Gen Scale
Context Window
200K (200,000 tokens)
Max Output Tokens
128K (128,000 tokens)
Knowledge Cutoff
March 2026
Supported Modalities
text, code, reasoning
License & Access
Z.AI Terms of Service

Benchmark Evaluations

Mmlu Pro
86.8
Humaneval
96
Swe Bench Verified
74.5
Chatbot Arena ELO
1465

Deep Architectural Overview

GLM-5.1 enables a closed loop from planning and execution to iterative refinement and final delivery. Built with multi-turn SFT and reinforcement learning with process-quality evaluation, it maintains remarkable goal consistency over 8-hour autonomous runs.

Strengths & Considerations

Core Strengths
  • Sustained 8-hour autonomous single-task execution
  • Comprehensive alignment with Claude Opus 4.6
  • 128k maximum generation output tokens
Known Limitations
  • Succeeded by GLM-5.2 for 1M token context windows

Token & API Pricing

Input Tokens (1M)$1.40
Output Tokens (1M)$4.40
Cached Input (1M)$0.2600
Pricing is verified directly against Z.AI's developer documentation and API rate sheets.

Similar & Alternative Models

Explore other frontier models from Z.AI and comparable reasoning engines.

Browse all models
Z.AI

Ultra-low latency version of GLM-5.3-Flash engineered for high-concurrency real-time IDE code completion and interactive agents.

200K ctx$0.37 / 1M tok (Ultra-Low Latency MoE)
Z.AI

The first native multimodal model in the GLM-5 series: 320B parameters (18B active) combining linear and sparse attention for low-cost visual coding.

200K ctx$0.15 / 1M tok (320B MoE Visual Coding)
Z.AI

Z.AI’s premier flagship model, delivering a 50% performance gain on Code Bench and matching Claude Mythos 5 in cybersecurity and vulnerability discovery.

1M ctx$1.40 / 1M tok (Flagship Coding & Cyber)
Z.AI

Flagship model built for project-scale context, supporting truly usable 1M-token context with 128k output.

1M ctx$1.40 / 1M tok (1M Lossless Context)