Introducing GLM-4.5Z.AIReleased July 28, 2025

GLM-4.5

Native agentic LLM delivering doubled parameter efficiency and seamless one-click compatibility with the Claude Code CLI.

reasoning-llmProprietary API$0.60 / 1M tok (Claude Code Compatible)Context: 128K (128,000 tokens)Arena ELO: 1310

Technical Specifications

Architecture Type
Native Agentic Transformer
Total Parameters
Frontier Scale
Context Window
128K (128,000 tokens)
Max Output Tokens
32.768K (32,768 tokens)
Knowledge Cutoff
June 2025
Supported Modalities
text, code, reasoning
License & Access
Z.AI Terms of Service

Benchmark Evaluations

Mmlu Pro
74
Humaneval
86.5
Swe Bench Lite
41.2
Chatbot Arena ELO
1310

Deep Architectural Overview

GLM-4.5 was engineered from the ground up for agentic execution. It became widely adopted as an economical drop-in backend for coding tools like Claude Code, Cline, and OpenCode with $0.60/$2.20 pricing.

Strengths & Considerations

Core Strengths
  • Direct compatibility with Claude Code and agent CLI frameworks
  • Doubled parameter efficiency and fast code generation
  • Prompt caching hit rate at $0.11 / 1M tokens
Known Limitations
  • Succeeded by GLM-4.6 and GLM-4.7 for larger multi-file repos

Token & API Pricing

Input Tokens (1M)$0.60
Output Tokens (1M)$2.20
Cached Input (1M)$0.1100
Pricing is verified directly against Z.AI's developer documentation and API rate sheets.

Similar & Alternative Models

Explore other frontier models from Z.AI and comparable reasoning engines.

Browse all models
Z.AI

Ultra-low latency version of GLM-5.3-Flash engineered for high-concurrency real-time IDE code completion and interactive agents.

200K ctx$0.37 / 1M tok (Ultra-Low Latency MoE)
Z.AI

The first native multimodal model in the GLM-5 series: 320B parameters (18B active) combining linear and sparse attention for low-cost visual coding.

200K ctx$0.15 / 1M tok (320B MoE Visual Coding)
Z.AI

Z.AI’s premier flagship model, delivering a 50% performance gain on Code Bench and matching Claude Mythos 5 in cybersecurity and vulnerability discovery.

1M ctx$1.40 / 1M tok (Flagship Coding & Cyber)
Z.AI

Flagship model built for project-scale context, supporting truly usable 1M-token context with 128k output.

1M ctx$1.40 / 1M tok (1M Lossless Context)