Cost-efficient ByteDance Seed 2.0 model for production chat, analysis, and structured generation
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Prior MiniMax coding model for agent workflows, office edits, and automation
General GLM flagship for coding, analysis, and tool-heavy engineering workflows
High-end Claude for difficult coding, planning, and slower expert reasoning
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Open-weight Qwen coding model for agents, repository edits, and multi-turn tool use
StepFun flash lane for quick multimodal reasoning and coding assistance
Budget GLM lane for fast coding help, routing, and everyday automation
Flagship model for demanding analysis, coding, and production agent workflows
Earlier Kimi frontier model for long-context agents, coding, and multimodal work
Earlier MiniMax agent model for practical coding and productivity tasks
Mature GLM model for dependable coding, reasoning, and structured agent tasks
New Gemini flash lane bringing frontier-style multimodal reasoning to cheaper runs
Small Nemotron 3 MoE for efficient coding, math, and long-context agents
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
Higher-accuracy GPT-5.2 variant for tougher reasoning and review workflows
Code-specialist GPT for repository edits, reviews, and long-running software agents
Reliable GPT generation for broad coding, writing, and tool-assisted product work
Mistral's coding-agent model for repository work, terminal tasks, and software fixes
GLM vision model for visual reasoning, documents, and multimodal agents
Multimodal reasoning model for visual analysis, planning, and tool use
Hybrid-reasoning DeepSeek model with thinking and non-thinking modes, sparse attention, and tool-use
DeepSeek chat model for instruction following, coding, and analysis
Flagship Claude model for deep reasoning, coding, and long-horizon agents
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Codex GPT for repository edits, code review, and practical software agents
Chat-tuned GPT-5.1 for polished assistants, writing, and product conversations
Sharper GPT-5 generation for coding, product work, and tool-assisted tasks
Thinking Kimi model for slower research passes, planning, and hard technical questions
Safety model for policy screening, moderation, and risk-aware routing workflows
Efficient open MiniMax model built for coding agents and tool-heavy workflows
Fast Claude lane for lightweight agents, office tasks, and responsive chat
ByteDance Seed model for long-context reasoning, instruction following, and tool-assisted tasks
Higher-accuracy GPT-5 tier for tough analysis, coding reviews, and planning
Late GLM-4 workhorse for coding agents, reasoning, and structured tasks
Balanced Claude model for coding, analysis, agent workflows, and cost control
Flagship Qwen3 model for coding agents, complex reasoning, and tool use
Qwen vision-language instruct model for visual reasoning, documents, and agent tasks
Qwen vision-language thinking model for visual reasoning, documents, and agent tasks
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Qwen instruction model for multilingual chat, reasoning, and tool use
Efficient Qwen thinking model for local reasoning, math, and coding agents
Low-latency ByteDance Seed model for high-throughput chat, extraction, and lightweight tool use
Tiny GPT-5 lane for routing, extraction, classification, and bulk jobs
Small GPT-5 for responsive agents, coding help, and everyday automation
Original GPT-5 workhorse for reasoning, coding, writing, and tool workflows
Flagship Claude model for deep reasoning, coding, and long-horizon agents
Open GPT reasoning model for self-hosted agents and controllable deployments
Open GPT reasoning model for self-hosted agents and controllable deployments
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Seed 2.0 Litebytedance-seed/seed-2.0-lite | 256K | $0.089 | $0.534 | 2026-02-14 | |||
| MiniMax-M2.5minimax/MiniMax-M2.5 | 204.8K | $0.3 | $1.2 | 2026-02-12 | |||
| GLM-5zhipuai/glm-5 | 204.8K | $1 | $3.2 | 2026-02-12 | |||
| Claude Opus 4.6anthropic/claude-opus-4-6 | 1M | $5 | $25 | 2026-02-05 | |||
| GPT-5.3 Codexopenai/gpt-5.3-codex | 400K | $1.75 | $14 | 2026-02-05 | |||
| Qwen3 Coder Nextalibaba/qwen3-coder-next | 262.144K | $0.108 | $0.675 | 2026-02-03 | |||
| Step 3.5 Flashstepfun/step-3.5-flash | 256K | $0.1 | $0.3 | 2026-01-29 | |||
| GLM-4.7-Flashzhipuai/glm-4.7-flash | 200K | $0.06 | $0.4 | 2026-01-19 | |||
| Solar Pro 3upstage/solar-pro3 | 131.072K | $0.15 | $0.6 | 2026-01 | |||
| Kimi K2.5moonshotai/kimi-k2.5 | 262.144K | $0.3 | $1.9 | 2026-01 | |||
| MiniMax-M2.1minimax/MiniMax-M2.1 | 204.8K | $0.3 | $1.2 | 2025-12-23 | |||
| GLM-4.7zhipuai/glm-4.7 | 204.8K | $0.6 | $2.2 | 2025-12-22 | |||
| Gemini 3 Flash Previewgoogle/gemini-3-flash-preview | 1.04858M | $0.5 | $3 | 2025-12-17 | |||
| Nemotron 3 Nano 30B A3Bnvidia/nemotron-3-nano-30b-a3b | 262.144K | $0.05 | $0.2 | 2025-12-15 | |||
| GPT-5.2 Chatopenai/gpt-5.2-chat-latest | 128K | $1.75 | $14 | 2025-12-11 | |||
| GPT-5.2 Proopenai/gpt-5.2-pro | 400K | $21 | $168 | 2025-12-11 | |||
| GPT-5.2 Codexopenai/gpt-5.2-codex | 400K | $0.14 | $1.14 | 2025-12-11 | |||
| GPT-5.2openai/gpt-5.2 | 400K | $1.75 | $14 | 2025-12-11 | |||
| Devstral 2mistral/devstral-2512 | 262.144K | $0.4 | $2 | 2025-12-09 | |||
| GLM-4.6Vzhipuai/glm-4.6v | 128K | $0.3 | $0.9 | 2025-12-08 | |||
| Nova 2 Liteamazon/nova-2-lite | 1M | $0.3 | $2.5 | 2025-12-02 | |||
| DeepSeek V3.2deepseek/deepseek-v3.2 | 128K | $0.18 | $0.35 | 2025-12-01 | |||
| DeepSeek Chatdeepseek/deepseek-chat | 1M | $0.147 | $0.295 | 2025-12-01 | |||
| Claude Opus 4.5 (latest)anthropic/claude-opus-4-5 | 200K | $5 | $25 | 2025-11-24 | |||
| GPT-5.1 Codex Maxopenai/gpt-5.1-codex-max | 400K | $1.1 | $9 | 2025-11-13 | |||
| GPT-5.1 Codex miniopenai/gpt-5.1-codex-mini | 400K | $0.22 | $1.8 | 2025-11-13 | |||
| GPT-5.1 Codexopenai/gpt-5.1-codex | 400K | $1.07 | $8.5 | 2025-11-13 | |||
| GPT-5.1 Chatopenai/gpt-5.1-chat-latest | 128K | $1.25 | $10 | 2025-11-13 | |||
| GPT-5.1openai/gpt-5.1 | 400K | $1.25 | $10 | 2025-11-13 | |||
| Kimi K2 Thinkingmoonshotai/kimi-k2-thinking | 262.144K | $0.4 | $2.5 | 2025-11-06 | |||
| GPT OSS Safeguard 20Bopenai/gpt-oss-safeguard-20b | 131.072K | $0.07 | $0.2 | 2025-10-29 | |||
| MiniMax-M2minimax/MiniMax-M2 | 204.8K | $0.3 | $1.2 | 2025-10-27 | |||
| Claude Haiku 4.5 (latest)anthropic/claude-haiku-4-5 | 200K | $1 | $5 | 2025-10-15 | |||
| Seed 1.6bytedance-seed/seed-1-6 | 256K | $0.119 | $1.187 | 2025-10-15 | |||
| GPT-5 Proopenai/gpt-5-pro | 400K | $15 | $120 | 2025-10-06 | |||
| GLM-4.6zhipuai/glm-4.6 | 204.8K | $0.6 | $2.2 | 2025-09-30 | |||
| Claude Sonnet 4.5 (latest)anthropic/claude-sonnet-4-5 | 200K | $3 | $15 | 2025-09-29 | |||
| Qwen3 Maxalibaba/qwen3-max | 262.144K | $1.2 | $6 | 2025-09-23 | |||
| Qwen3 VL 235B A22B Instructalibaba/qwen3-vl-235b-a22b-instruct | 131.072K | $0.2 | $0.88 | 2025-09-23 | |||
| Qwen3 VL 235B A22B Thinkingalibaba/qwen3-vl-235b-a22b-thinking | 131.072K | $0.4 | $4 | 2025-09-23 | |||
| GPT-5-Codexopenai/gpt-5-codex | 400K | $1.1 | $9 | 2025-09-15 | |||
| Qwen3-Next 80B-A3B Instructalibaba/qwen3-next-80b-a3b-instruct | 131.072K | $0.5 | $2 | 2025-09 | |||
| Qwen3-Next 80B-A3B (Thinking)alibaba/qwen3-next-80b-a3b-thinking | 131.072K | $0.5 | $6 | 2025-09 | |||
| Seed 1.6 Flashbytedance-seed/seed-1-6-flash | 256K | $0.022 | $0.223 | 2025-08-28 | |||
| GPT-5 Nanoopenai/gpt-5-nano | 400K | $0.05 | $0.4 | 2025-08-07 | |||
| GPT-5 Miniopenai/gpt-5-mini | 400K | $0.25 | $2 | 2025-08-07 | |||
| GPT-5openai/gpt-5 | 400K | $1.25 | $10 | 2025-08-07 | |||
| Claude Opus 4.1 (latest)anthropic/claude-opus-4-1 | 200K | $15 | $75 | 2025-08-05 | |||
| GPT OSS 120Bopenai/gpt-oss-120b | 131.072K | $0.03 | $0.17 | 2025-08-05 | |||
| GPT OSS 20Bopenai/gpt-oss-20b | 131.072K | $0.02 | $0.1 | 2025-08-05 |