Open MiniMax flagship for coding agents, office automation, and complex environments
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Cheapest GPT-5.4 lane for simple routing, extraction, and bulk automation
Strong small GPT for coding subagents, quick tool use, and high-volume work
Efficient Mistral model for fast chat, extraction, and production assistants
Faster GLM-5 lane for coding agents that need lower latency
Fast Mistral production model for chat, extraction, and cost-sensitive agents
Nemotron middle tier for collaborative agents and high-volume reasoning workloads
Grok model for agentic tool use, reasoning, coding, and live assistance
Reasoning Grok for document-heavy analysis and long-horizon tool use
Agent-ready GPT for coding and computer-use workflows at a lower cost
More exact GPT-5.4 tier for demanding professional reasoning and agent tasks
Low-latency Gemini model for high-volume multimodal and agent workloads
Advanced Gemini model for complex reasoning, coding, and multimodal analysis
Reasoning-first Gemini preview for agentic coding and complex problem solving
Claude workhorse for coding agents, careful analysis, and production cost control
Lightweight ByteDance Seed 2.0 model for low-latency multimodal reasoning and high-volume tasks
ByteDance Seed coding model for multimodal software engineering and long-running agents
General GLM flagship for coding, analysis, and tool-heavy engineering workflows
Prior MiniMax coding model for agent workflows, office edits, and automation
Coding-optimized GPT model for repository edits, reviews, and agentic software work
High-end Claude for difficult coding, planning, and slower expert reasoning
Open-weight Qwen coding model for agents, repository edits, and multi-turn tool use
StepFun flash lane for quick multimodal reasoning and coding assistance
Budget GLM lane for fast coding help, routing, and everyday automation
Earlier Kimi frontier model for long-context agents, coding, and multimodal work
Earlier MiniMax agent model for practical coding and productivity tasks
Mature GLM model for dependable coding, reasoning, and structured agent tasks
New Gemini flash lane bringing frontier-style multimodal reasoning to cheaper runs
Small Nemotron 3 MoE for efficient coding, math, and long-context agents
Code-specialist GPT for repository edits, reviews, and long-running software agents
Higher-accuracy GPT-5.2 variant for tougher reasoning and review workflows
Reliable GPT generation for broad coding, writing, and tool-assisted product work
Mistral's coding-agent model for repository work, terminal tasks, and software fixes
Mistral coding agent model for repository tasks and software engineering workflows
DeepSeek reasoning model for multi-step analysis, math, coding, and tools
DeepSeek chat model for instruction following, coding, and analysis
Flagship Claude model for deep reasoning, coding, and long-horizon agents
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Sharper GPT-5 generation for coding, product work, and tool-assisted tasks
Codex GPT for repository edits, code review, and practical software agents
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Thinking Kimi model for slower research passes, planning, and hard technical questions
Flagship Claude model for deep reasoning, coding, and long-horizon agents
Efficient open MiniMax model built for coding agents and tool-heavy workflows
Higher-accuracy GPT-5 tier for tough analysis, coding reviews, and planning
Late GLM-4 workhorse for coding agents, reasoning, and structured tasks
Flagship Qwen3 model for coding agents, complex reasoning, and tool use
Small GPT-5 for responsive agents, coding help, and everyday automation
Tiny GPT-5 lane for routing, extraction, classification, and bulk jobs
Original GPT-5 workhorse for reasoning, coding, writing, and tool workflows
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| MiniMax-M2.7minimax/MiniMax-M2.7 | 204.8K | $0.3 | $1.2 | 2026-03-18 | |||
| GPT-5.4 nanoopenai/gpt-5.4-nano | 400K | $0.2 | $1.25 | 2026-03-17 | |||
| GPT-5.4 miniopenai/gpt-5.4-mini | 400K | $0.75 | $4.5 | 2026-03-17 | |||
| Mistral Small (latest)mistral/mistral-small-latest | 256K | $0.15 | $0.6 | 2026-03-16 | |||
| GLM-5-Turbozhipuai/glm-5-turbo | 200K | $0.9 | $3.7 | 2026-03-16 | |||
| Mistral Small 4mistral/mistral-small-2603 | 256K | $0.15 | $0.6 | 2026-03-16 | |||
| Nemotron 3 Super 120B A12Bnvidia/nemotron-3-super-120b-a12b | 262.144K | $0.2 | $0.8 | 2026-03-11 | |||
| Grok 4.20 (Non-Reasoning)xai/grok-4.20-0309-non-reasoning | 1M | $1.25 | $2.5 | 2026-03-09 | |||
| Grok 4.20 (Reasoning)xai/grok-4.20-0309-reasoning | 1M | $1.25 | $2.5 | 2026-03-09 | |||
| GPT-5.4openai/gpt-5.4 | 1.05M | $2.5 | $15 | 2026-03-05 | |||
| GPT-5.4 Proopenai/gpt-5.4-pro | 1.05M | $30 | $180 | 2026-03-05 | |||
| Gemini 3.1 Flash Lite Previewgoogle/gemini-3.1-flash-lite-preview | 1.04858M | $0.25 | $1.5 | 2026-03-03 | |||
| Gemini 3.1 Pro Preview Custom Toolsgoogle/gemini-3.1-pro-preview-customtools | 1.04858M | $2 | $12 | 2026-02-19 | |||
| Gemini 3.1 Pro Previewgoogle/gemini-3.1-pro-preview | 1.04858M | $2 | $12 | 2026-02-19 | |||
| Claude Sonnet 4.6anthropic/claude-sonnet-4-6 | 1M | $3 | $15 | 2026-02-17 | |||
| Seed 2.0 Minibytedance-seed/seed-2.0-mini | 256K | $0.03 | $0.297 | 2026-02-14 | |||
| Seed 2.0 Codebytedance-seed/seed-2.0-code | 262.144K | $0.4 | $2.4 | 2026-02-14 | |||
| GLM-5zhipuai/glm-5 | 204.8K | $1 | $3.2 | 2026-02-12 | |||
| MiniMax-M2.5minimax/MiniMax-M2.5 | 204.8K | $0.3 | $1.2 | 2026-02-12 | |||
| GPT-5.3 Codexopenai/gpt-5.3-codex | 400K | $1.75 | $14 | 2026-02-05 | |||
| Claude Opus 4.6anthropic/claude-opus-4-6 | 1M | $5 | $25 | 2026-02-05 | |||
| Qwen3 Coder Nextalibaba/qwen3-coder-next | 262.144K | $0.108 | $0.675 | 2026-02-03 | |||
| Step 3.5 Flashstepfun/step-3.5-flash | 256K | $0.1 | $0.3 | 2026-01-29 | |||
| GLM-4.7-Flashzhipuai/glm-4.7-flash | 200K | $0.06 | $0.4 | 2026-01-19 | |||
| Kimi K2.5moonshotai/kimi-k2.5 | 262.144K | $0.3 | $1.9 | 2026-01 | |||
| MiniMax-M2.1minimax/MiniMax-M2.1 | 204.8K | $0.3 | $1.2 | 2025-12-23 | |||
| GLM-4.7zhipuai/glm-4.7 | 204.8K | $0.6 | $2.2 | 2025-12-22 | |||
| Gemini 3 Flash Previewgoogle/gemini-3-flash-preview | 1.04858M | $0.5 | $3 | 2025-12-17 | |||
| Nemotron 3 Nano 30B A3Bnvidia/nemotron-3-nano-30b-a3b | 262.144K | $0.05 | $0.2 | 2025-12-15 | |||
| GPT-5.2 Codexopenai/gpt-5.2-codex | 400K | $0.14 | $1.14 | 2025-12-11 | |||
| GPT-5.2 Proopenai/gpt-5.2-pro | 400K | $21 | $168 | 2025-12-11 | |||
| GPT-5.2openai/gpt-5.2 | 400K | $1.75 | $14 | 2025-12-11 | |||
| Devstral 2mistral/devstral-2512 | 262.144K | $0.4 | $2 | 2025-12-09 | |||
| Devstral 2 (latest)mistral/devstral-medium-latest | 262.144K | $0.4 | $2 | 2025-12-02 | |||
| DeepSeek Reasonerdeepseek/deepseek-reasoner | 1M | $0.147 | $0.295 | 2025-12-01 | |||
| DeepSeek Chatdeepseek/deepseek-chat | 1M | $0.147 | $0.295 | 2025-12-01 | |||
| Claude Opus 4.5 (latest)anthropic/claude-opus-4-5 | 200K | $5 | $25 | 2025-11-24 | |||
| GPT-5.1 Codex miniopenai/gpt-5.1-codex-mini | 400K | $0.22 | $1.8 | 2025-11-13 | |||
| GPT-5.1openai/gpt-5.1 | 400K | $1.25 | $10 | 2025-11-13 | |||
| GPT-5.1 Codexopenai/gpt-5.1-codex | 400K | $1.07 | $8.5 | 2025-11-13 | |||
| GPT-5.1 Codex Maxopenai/gpt-5.1-codex-max | 400K | $1.1 | $9 | 2025-11-13 | |||
| Kimi K2 Thinkingmoonshotai/kimi-k2-thinking | 262.144K | $0.4 | $2.5 | 2025-11-06 | |||
| Claude Opus 4.5anthropic/claude-opus-4-5-20251101 | 200K | $5 | $25 | 2025-11-01 | |||
| MiniMax-M2minimax/MiniMax-M2 | 204.8K | $0.3 | $1.2 | 2025-10-27 | |||
| GPT-5 Proopenai/gpt-5-pro | 400K | $15 | $120 | 2025-10-06 | |||
| GLM-4.6zhipuai/glm-4.6 | 204.8K | $0.6 | $2.2 | 2025-09-30 | |||
| Qwen3 Maxalibaba/qwen3-max | 262.144K | $1.2 | $6 | 2025-09-23 | |||
| GPT-5 Miniopenai/gpt-5-mini | 400K | $0.25 | $2 | 2025-08-07 | |||
| GPT-5 Nanoopenai/gpt-5-nano | 400K | $0.05 | $0.4 | 2025-08-07 | |||
| GPT-5openai/gpt-5 | 400K | $1.25 | $10 | 2025-08-07 |