Top-performing open-weight model optimized for frontier reasoning, coding, and enterprise AI applications.
Flagship conversational model built for real-time knowledge exploration, sharp reasoning, and highly engaging AI interactions.
High-end coding agent built for complex software engineering, repository-scale tasks, and autonomous development workflows.
Fast, lightweight coding model designed for interactive development, rapid code iteration, and efficient coding agents.

Flagship model delivering premium reasoning, coding, multimodal understanding, and enterprise-grade performance.

Turbo model optimized for ultra-low latency, high throughput, and responsive AI experiences.

Flagship model delivering premium reasoning, coding, multimodal understanding, and enterprise-grade performance.
Agent-oriented model built for complex reasoning, tool use, and autonomous task execution.
Powerful coding model for programming, debugging, and AI developer workflows.
MiniMax M3 is a lightweight, state-of-the-art large language model optimized for coding, agentic workflows, and modern application development. With only 10 billion activated parameters, it delivers a major jump in real-world capability while maintaining exceptional latency, scalability, and cost efficiency.
Specialized coding model optimized for software development, code generation, debugging, refactoring, and developer workflows.
Advanced conversational AI model optimized for natural dialogue, knowledge exploration, reasoning, and interactive chat experiences.
Fast and cost-efficient multimodal model designed for high-throughput applications, real-time interactions, and everyday AI tasks.
DeepSeek V4 Pro is a state-of-the-art large language model combining efficient sparse attention, strong reasoning, and integrated agent capabilities for robust long-context understanding and versatile AI applications.
DeepSeek V4 Flash is a state-of-the-art large language model combining efficient sparse attention, strong reasoning, and integrated agent capabilities for robust long-context understanding and versatile AI applications.
DeepSeek V4 Flash is a state-of-the-art large language model combining efficient sparse attention, strong reasoning, and integrated agent capabilities for robust long-context understanding and versatile AI applications.
Processes text, image, audio, and video to text. Delivers strong agentic performance with high token efficiency and low inference cost.
Optimized for complex reasoning, software engineering, and long-horizon agentic tasks. Supports autonomous execution of extensive tool-calling workflows.

Flagship foundation model built for deep reasoning, complex problem-solving, and sophisticated agentic workflows.
Enhanced model for reasoning, coding, and productivity.
The latest Qwen reasoning model.
Versatile model for chat, and productivity workflows.

Self-improving research model designed for adaptive reasoning, exploration, and continuous learning workflows.

Professional-grade model built for advanced workloads, complex analysis, and enterprise AI applications.

Developer-focused model specialized in coding agents, repository understanding, and software engineering.

Ultra-efficient model focused on lightweight AI tasks, rapid inference, and large-scale deployment.

Small yet capable model designed for edge scenarios, automation, and cost-sensitive services.

Next-generation assistant model with improved instruction following and deeper contextual understanding.

High-speed model engineered for instant responses, real-time interaction, and massive request workloads.

Versatile foundation model providing reliable conversation, knowledge understanding, and content creation.
GLM-5.1 is Z.AI’s latest flagship model, featuring upgrades in two key areas: enhanced programming capabilities and more stable multi-step reasoning/execution. It demonstrates significant improvements in executing complex agent tasks while delivering more natural conversational experiences and superior front-end aesthetics.
MiniMax-M2.7 is a lightweight, state-of-the-art large language model optimized for coding, agentic workflows, and modern application development. With only 10 billion activated parameters, it delivers a major jump in real-world capability while maintaining exceptional latency, scalability, and cost efficiency.
Qwen3.5 represents a significant leap forward, integrating breakthroughs in multimodal learning, architectural efficiency, reinforcement learning scale, and global accessibility to empower developers and enterprises with unprecedented capability and efficiency.
Qwen3.5 represents a significant leap forward, integrating breakthroughs in multimodal learning, architectural efficiency, reinforcement learning scale, and global accessibility to empower developers and enterprises with unprecedented capability and efficiency.
Qwen3.5 represents a significant leap forward, integrating breakthroughs in multimodal learning, architectural efficiency, reinforcement learning scale, and global accessibility to empower developers and enterprises with unprecedented capability and efficiency.
Qwen3.5 represents a significant leap forward, integrating breakthroughs in multimodal learning, architectural efficiency, reinforcement learning scale, and global accessibility to empower developers and enterprises with unprecedented capability and efficiency.
MiniMax-M2.5 is a lightweight, state-of-the-art large language model optimized for coding, agentic workflows, and modern application development. With only 10 billion activated parameters, it delivers a major jump in real-world capability while maintaining exceptional latency, scalability, and cost efficiency.
GLM-5v Turbo is Z.AI’s latest flagship model, featuring upgrades in two key areas: enhanced programming capabilities and more stable multi-step reasoning/execution. It demonstrates significant improvements in executing complex agent tasks while delivering more natural conversational experiences and superior front-end aesthetics.
GLM-5 is Z.AI’s latest flagship model, featuring upgrades in two key areas: enhanced programming capabilities and more stable multi-step reasoning/execution. It demonstrates significant improvements in executing complex agent tasks while delivering more natural conversational experiences and superior front-end aesthetics.
Powerful model for long-context and intelligent workflows.
Flagship model for advanced reasoning, coding, and complex tasks.
Balanced model combining strong capability, speed, and efficiency.
Efficient model for everyday tasks and AI assistants.
Fast model optimized for instant responses and large-scale usage.
GLM-4.7 is Z.AI’s latest flagship model, featuring upgrades in two key areas: enhanced programming capabilities and more stable multi-step reasoning/execution. It demonstrates significant improvements in executing complex agent tasks while delivering more natural conversational experiences and superior front-end aesthetics.
DeepSeek V3.2 is a state-of-the-art large language model combining efficient sparse attention, strong reasoning, and integrated agent capabilities for robust long-context understanding and versatile AI applications.
High-intelligence model built for deep reasoning, ambitious problem-solving, and frontier AI workloads.
Dependable general-purpose model designed for practical workflows, grounded analysis, and production applications.