Ultra-low latency version of GLM-5.3-Flash engineered for high-concurrency real-time IDE code completion and interactive agents.
AI Models Directory
Comprehensive directory of 23 foundation models, LLMs, and multimodal engines from top research labs. Real pricing, benchmark scores, and architecture cards.
The first native multimodal model in the GLM-5 series: 320B parameters (18B active) combining linear and sparse attention for low-cost visual coding.
Z.AI’s premier flagship model, delivering a 50% performance gain on Code Bench and matching Claude Mythos 5 in cybersecurity and vulnerability discovery.
Flagship model built for project-scale context, supporting truly usable 1M-token context with 128k output.
Engineered for long-horizon tasks, able to work independently for up to 8 hours in a single run, aligned with Claude Opus 4.6.
Fifth-generation foundation model shifting from coding to complex systems engineering, benchmarked against Claude Opus 4.5.
Compact optical character recognition model combining CogViT with GLM-0.5B for fast, highly accurate document and table extraction.
Free-tier version of GLM-4.7 delivering fast inference, 200k context, and strong coding for high-frequency applications.
State-of-the-art image generation model combining autoregressive semantic understanding with diffusion decoding for accurate text rendering.
Optimized agentic coding model setting open-source SOTA performance on major coding and reasoning benchmarks with 200k context.
Multimodal mobile automation framework that understands smartphone screen content and executes real user actions through ADB across 50+ mainstream apps.
High-accuracy automatic speech recognition model delivering a Character Error Rate as low as 0.0717 with user-defined vocabulary support.
Multimodal vision model with native function calling and controllable thinking mode switch over 128k tokens.
Flagship coding model leading Chinese programming benchmarks with 200k context and 64k maximum output tokens.
A 100B-scale open-weights vision reasoning model supporting video understanding, visual grounding, and GUI agents.
Completely free, high-speed lightweight model supporting 200k context for developers and hobbyists.
Cost-effective high-performance variant in the GLM-4.5 family, priced at only $0.20 per million input tokens.
Native agentic LLM delivering doubled parameter efficiency and seamless one-click compatibility with the Claude Code CLI.
Advanced video generation model featuring start and end frame synthesis, improved physical realism simulation, and high visual stability.
A 32B parameter model delivering high intelligence at unmatched cost efficiency with flat $0.10/1M input and output pricing.
High-quality diffusion image generation model producing rich detail, photographic textures, and artistic styling at $0.01 per image.
End-to-end speech conversational model capable of understanding and generating emotional, realtime human-like speech without intermediate ASR/TTS.
The breakthrough foundation model by Zhipu AI / Z.AI featuring 128k context, strong bilingual comprehension, and vision capabilities.