Frontier Model Intelligence

AI Models Directory

Comprehensive directory of 23 foundation models, LLMs, and multimodal engines from top research labs. Real pricing, benchmark scores, and architecture cards.

Z.AI

Ultra-low latency version of GLM-5.3-Flash engineered for high-concurrency real-time IDE code completion and interactive agents.

Context :200K ctx
Pricing :$0.37 / 1M tok (Ultra-Low Latency MoE)
Type :multimodal-realtime
Z.AI

The first native multimodal model in the GLM-5 series: 320B parameters (18B active) combining linear and sparse attention for low-cost visual coding.

Context :200K ctx
Pricing :$0.15 / 1M tok (320B MoE Visual Coding)
Type :multimodal-realtime
Z.AI

Z.AI’s premier flagship model, delivering a 50% performance gain on Code Bench and matching Claude Mythos 5 in cybersecurity and vulnerability discovery.

Context :1M ctx
Pricing :$1.40 / 1M tok (Flagship Coding & Cyber)
Type :reasoning-llm
Z.AI

Flagship model built for project-scale context, supporting truly usable 1M-token context with 128k output.

Context :1M ctx
Pricing :$1.40 / 1M tok (1M Lossless Context)
Type :reasoning-llm
Z.AI

Engineered for long-horizon tasks, able to work independently for up to 8 hours in a single run, aligned with Claude Opus 4.6.

Context :200K ctx
Pricing :$1.40 / 1M tok (8h Autonomous Agent)
Type :reasoning-llm
Z.AI

Fifth-generation foundation model shifting from coding to complex systems engineering, benchmarked against Claude Opus 4.5.

Context :200K ctx
Pricing :$1.00 / 1M tok (DeepSeek Sparse Attention)
Type :reasoning-llm
Z.AI

Compact optical character recognition model combining CogViT with GLM-0.5B for fast, highly accurate document and table extraction.

Context :33K ctx
Pricing :$0.03 / 1M tok (CogViT + GLM-0.5B)
Type :multimodal-realtime
Z.AI

Free-tier version of GLM-4.7 delivering fast inference, 200k context, and strong coding for high-frequency applications.

Context :200K ctx
Pricing :Free API (200k Context 4.7)
Type :text-to-text
Z.AI

State-of-the-art image generation model combining autoregressive semantic understanding with diffusion decoding for accurate text rendering.

Context :4K ctx
Pricing :$0.015 / image (Text Rendering SOTA)
Type :text-to-image
Z.AI

Optimized agentic coding model setting open-source SOTA performance on major coding and reasoning benchmarks with 200k context.

Context :200K ctx
Pricing :$0.60 / 1M tok (Agentic Coding SOTA)
Type :reasoning-llm
Z.AI

Multimodal mobile automation framework that understands smartphone screen content and executes real user actions through ADB across 50+ mainstream apps.

Context :64K ctx
Pricing :Mobile OS Automation Agent
Type :autonomous-agent
Z.AI

High-accuracy automatic speech recognition model delivering a Character Error Rate as low as 0.0717 with user-defined vocabulary support.

Context :33K ctx
Pricing :$0.0024 / min audio (0.0717 CER)
Type :speech-to-text
Z.AI

Multimodal vision model with native function calling and controllable thinking mode switch over 128k tokens.

Context :128K ctx
Pricing :$0.30 / 1M tok (Native Tool Calling Vision)
Type :multimodal-realtime
Z.AI

Flagship coding model leading Chinese programming benchmarks with 200k context and 64k maximum output tokens.

Context :200K ctx
Pricing :$0.60 / 1M tok (China Coding Leader)
Type :reasoning-llm
Z.AI

A 100B-scale open-weights vision reasoning model supporting video understanding, visual grounding, and GUI agents.

Context :64K ctx
Pricing :$0.60 / 1M tok (100B Open Vision)
Type :multimodal-realtime
Z.AI

Completely free, high-speed lightweight model supporting 200k context for developers and hobbyists.

Context :200K ctx
Pricing :Free API (200k Context Flash)
Type :text-to-text
Z.AI

Cost-effective high-performance variant in the GLM-4.5 family, priced at only $0.20 per million input tokens.

Context :128K ctx
Pricing :$0.20 / 1M tok (Cost-Effective Air)
Type :text-to-text
Z.AI

Native agentic LLM delivering doubled parameter efficiency and seamless one-click compatibility with the Claude Code CLI.

Context :128K ctx
Pricing :$0.60 / 1M tok (Claude Code Compatible)
Type :reasoning-llm
Z.AI

Advanced video generation model featuring start and end frame synthesis, improved physical realism simulation, and high visual stability.

Context :4K ctx
Pricing :$0.20 / video (Keyframe Synthesis)
Type :text-to-video
Z.AI

A 32B parameter model delivering high intelligence at unmatched cost efficiency with flat $0.10/1M input and output pricing.

Context :128K ctx
Pricing :$0.10 / 1M tok (Symmetric Price)
Type :text-to-text
Z.AI

High-quality diffusion image generation model producing rich detail, photographic textures, and artistic styling at $0.01 per image.

Context :4K ctx
Pricing :$0.010 / image
Type :text-to-image
Z.AI

End-to-end speech conversational model capable of understanding and generating emotional, realtime human-like speech without intermediate ASR/TTS.

Context :33K ctx
Pricing :Free Open Weights / End-to-End Voice
Type :speech-to-speech
Z.AI

The breakthrough foundation model by Zhipu AI / Z.AI featuring 128k context, strong bilingual comprehension, and vision capabilities.

Context :128K ctx
Pricing :$1.40 / 1M tok (Flagship Foundation)
Type :text-to-text