Default frontier GPT for coding, computer use, research, and knowledge work
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution
Qwen vision-language model for visual reasoning, documents, and agent tasks
Open MiMo model for multimodal coding agents and long-context automation
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
Flagship Qwen model for complex reasoning, coding, and agentic workflows
Tencent Hy reasoning model for coding, instruction following, and agent tasks
xAI's default Grok for chat, coding, agentic tools, and lower hallucination risk
Fast Grok coding model tuned for agentic engineering and iterative edits
Muse Spark is a natively multimodal reasoning model with support for tool-use, visual chain of thought, and multi-agent orchestration.
Largest Gemma 4 instruction model for open, self-hosted chat and reasoning
StepFun flash model for efficient multimodal reasoning, coding, and tool use
Open Gemma instruction model for efficient chat and self-hosted deployments
Earlier MiMo Pro model for multimodal agents, reasoning, and code tasks
Strong small GPT for coding subagents, quick tool use, and high-volume work
Cheapest GPT-5.4 lane for simple routing, extraction, and bulk automation
Nemotron middle tier for collaborative agents and high-volume reasoning workloads
More exact GPT-5.4 tier for demanding professional reasoning and agent tasks
Agent-ready GPT for coding and computer-use workflows at a lower cost
Qwen vision-language model for visual reasoning, documents, and agent tasks
Qwen instruction model for multilingual chat, reasoning, and tool use
Qwen vision-language model for visual reasoning, documents, and agent tasks
Reasoning-first Gemini preview for agentic coding and complex problem solving
Advanced Gemini model for complex reasoning, coding, and multimodal analysis
Claude workhorse for coding agents, careful analysis, and production cost control
Qwen vision-language model for visual reasoning, documents, and agent tasks
Cost-efficient ByteDance Seed 2.0 model for production chat, analysis, and structured generation
ByteDance Seed coding model for multimodal software engineering and long-running agents
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Open-weight Qwen coding model for agents, repository edits, and multi-turn tool use
StepFun flash lane for quick multimodal reasoning and coding assistance
Flagship model for demanding analysis, coding, and production agent workflows
Earlier Kimi frontier model for long-context agents, coding, and multimodal work
New Gemini flash lane bringing frontier-style multimodal reasoning to cheaper runs
Small Nemotron 3 MoE for efficient coding, math, and long-context agents
Reliable GPT generation for broad coding, writing, and tool-assisted product work
Higher-accuracy GPT-5.2 variant for tougher reasoning and review workflows
Code-specialist GPT for repository edits, reviews, and long-running software agents
Hybrid-reasoning DeepSeek model with thinking and non-thinking modes, sparse attention, and tool-use
DeepSeek reasoning model for multi-step analysis, math, coding, and tools
Codex GPT for repository edits, code review, and practical software agents
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Sharper GPT-5 generation for coding, product work, and tool-assisted tasks
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Thinking Kimi model for slower research passes, planning, and hard technical questions
Safety model for policy screening, moderation, and risk-aware routing workflows
Fast Claude model for responsive assistance, classification, and lightweight agents
Higher-accuracy GPT-5 tier for tough analysis, coding reviews, and planning
Balanced Claude model for coding, analysis, agent workflows, and cost control
Flagship Qwen3 model for coding agents, complex reasoning, and tool use
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| GPT-5.5openai/gpt-5.5 | 1.05M | $5 | $30 | 2026-04-23 | |||
| MiMo-V2.5-Proxiaomi/mimo-v2.5-pro | 1.04858M | $0.435 | $0.87 | 2026-04-22 | |||
| Qwen3.6 27Balibaba/qwen3.6-27b | 262.144K | $0.6 | $3.6 | 2026-04-22 | |||
| MiMo-V2.5xiaomi/mimo-v2.5 | 1.04858M | $0.14 | $0.28 | 2026-04-22 | |||
| Kimi K2.6moonshotai/kimi-k2.6 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| Qwen3.6 Max Previewalibaba/qwen3.6-max-preview | 262.144K | $1.3 | $7.8 | 2026-04-20 | |||
| Hy3 previewtencent/hy3-preview | 256K | $0.066 | $0.26 | 2026-04-20 | |||
| Grok 4.3xai/grok-4.3 | 1M | $1.25 | $2.5 | 2026-04-17 | |||
| Grok Build 0.1xai/grok-build-0.1 | 256K | $1 | $2 | 2026-04-16 | |||
| Muse Spark 1.1meta/muse-spark-1.1 | 1.04858M | $1.25 | $4.25 | 2026-04-08 | |||
| Gemma 4 31B ITgoogle/gemma-4-31b-it | 262.144K | $0.09 | $0.34 | 2026-04-02 | |||
| Step 3.5 Flash 2603stepfun/step-3.5-flash-2603 | 256K | $0.1 | $0.3 | 2026-04-02 | |||
| Gemma 4 26B A4B ITgoogle/gemma-4-26b-a4b-it | 262.144K | $0.042 | $0.22 | 2026-04-02 | |||
| MiMo-V2-Proxiaomi/mimo-v2-pro | 1.04858M | $0.435 | $0.87 | 2026-03-18 | |||
| GPT-5.4 miniopenai/gpt-5.4-mini | 400K | $0.75 | $4.5 | 2026-03-17 | |||
| GPT-5.4 nanoopenai/gpt-5.4-nano | 400K | $0.2 | $1.25 | 2026-03-17 | |||
| Nemotron 3 Super 120B A12Bnvidia/nemotron-3-super-120b-a12b | 262.144K | $0.2 | $0.8 | 2026-03-11 | |||
| GPT-5.4 Proopenai/gpt-5.4-pro | 1.05M | $30 | $180 | 2026-03-05 | |||
| GPT-5.4openai/gpt-5.4 | 1.05M | $2.5 | $15 | 2026-03-05 | |||
| Qwen3.5 Flashalibaba/qwen3.5-flash | 1M | $0.029 | $0.287 | 2026-02-23 | |||
| Qwen3.5 9Balibaba/qwen3.5-9b | 262.144K | $0.04 | $0.15 | 2026-02-23 | |||
| Qwen3.5 27Balibaba/qwen3.5-27b | 262.144K | $0.3 | $2.4 | 2026-02-23 | |||
| Gemini 3.1 Pro Previewgoogle/gemini-3.1-pro-preview | 1.04858M | $2 | $12 | 2026-02-19 | |||
| Gemini 3.1 Pro Preview Custom Toolsgoogle/gemini-3.1-pro-preview-customtools | 1.04858M | $2 | $12 | 2026-02-19 | |||
| Claude Sonnet 4.6anthropic/claude-sonnet-4-6 | 1M | $3 | $15 | 2026-02-17 | |||
| Qwen3.5 Plusalibaba/qwen3.5-plus | 1M | $0.4 | $2.4 | 2026-02-16 | |||
| Seed 2.0 Litebytedance-seed/seed-2.0-lite | 256K | $0.089 | $0.534 | 2026-02-14 | |||
| Seed 2.0 Codebytedance-seed/seed-2.0-code | 262.144K | $0.4 | $2.4 | 2026-02-14 | |||
| GPT-5.3 Codexopenai/gpt-5.3-codex | 400K | $1.75 | $14 | 2026-02-05 | |||
| Qwen3 Coder Nextalibaba/qwen3-coder-next | 262.144K | $0.108 | $0.675 | 2026-02-03 | |||
| Step 3.5 Flashstepfun/step-3.5-flash | 256K | $0.1 | $0.3 | 2026-01-29 | |||
| Solar Pro 3upstage/solar-pro3 | 131.072K | $0.15 | $0.6 | 2026-01 | |||
| Kimi K2.5moonshotai/kimi-k2.5 | 262.144K | $0.3 | $1.9 | 2026-01 | |||
| Gemini 3 Flash Previewgoogle/gemini-3-flash-preview | 1.04858M | $0.5 | $3 | 2025-12-17 | |||
| Nemotron 3 Nano 30B A3Bnvidia/nemotron-3-nano-30b-a3b | 262.144K | $0.05 | $0.2 | 2025-12-15 | |||
| GPT-5.2openai/gpt-5.2 | 400K | $1.75 | $14 | 2025-12-11 | |||
| GPT-5.2 Proopenai/gpt-5.2-pro | 400K | $21 | $168 | 2025-12-11 | |||
| GPT-5.2 Codexopenai/gpt-5.2-codex | 400K | $0.14 | $1.14 | 2025-12-11 | |||
| DeepSeek V3.2deepseek/deepseek-v3.2 | 128K | $0.18 | $0.35 | 2025-12-01 | |||
| DeepSeek Reasonerdeepseek/deepseek-reasoner | 1M | $0.147 | $0.295 | 2025-12-01 | |||
| GPT-5.1 Codexopenai/gpt-5.1-codex | 400K | $1.07 | $8.5 | 2025-11-13 | |||
| GPT-5.1 Codex miniopenai/gpt-5.1-codex-mini | 400K | $0.22 | $1.8 | 2025-11-13 | |||
| GPT-5.1openai/gpt-5.1 | 400K | $1.25 | $10 | 2025-11-13 | |||
| GPT-5.1 Codex Maxopenai/gpt-5.1-codex-max | 400K | $1.1 | $9 | 2025-11-13 | |||
| Kimi K2 Thinkingmoonshotai/kimi-k2-thinking | 262.144K | $0.4 | $2.5 | 2025-11-06 | |||
| GPT OSS Safeguard 20Bopenai/gpt-oss-safeguard-20b | 131.072K | $0.07 | $0.2 | 2025-10-29 | |||
| Claude Haiku 4.5anthropic/claude-haiku-4-5-20251001 | 200K | $1 | $5 | 2025-10-15 | |||
| GPT-5 Proopenai/gpt-5-pro | 400K | $15 | $120 | 2025-10-06 | |||
| Claude Sonnet 4.5anthropic/claude-sonnet-4-5-20250929 | 200K | $3 | $15 | 2025-09-29 | |||
| Qwen3 Maxalibaba/qwen3-max | 262.144K | $1.2 | $6 | 2025-09-23 |