Mature GLM model for dependable coding, reasoning, and structured agent tasks
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
New Gemini flash lane bringing frontier-style multimodal reasoning to cheaper runs
MiMo flash model for fast multimodal assistance and agent workflows
Small Nemotron 3 MoE for efficient coding, math, and long-context agents
Higher-accuracy GPT-5.2 variant for tougher reasoning and review workflows
Code-specialist GPT for repository edits, reviews, and long-running software agents
Reliable GPT generation for broad coding, writing, and tool-assisted product work
Mistral's coding-agent model for repository work, terminal tasks, and software fixes
Compact multimodal coding model for repository exploration, file editing, and software agents
Compact multimodal Mistral model for local assistants, edge agents, and efficient tool use
Mistral's largest general model for enterprise agents, coding, and multilingual reasoning
Mistral coding agent model for repository tasks and software engineering workflows
Compact open vision-language model for edge deployment, instruction following, and tool use
Compact open vision-language model for edge deployment, instruction following, and tool use
Open vision-language model for efficient local deployment, instruction following, and tool use
Multimodal reasoning model for visual analysis, planning, and tool use
DeepSeek reasoning model for multi-step analysis, math, coding, and tools
DeepSeek chat model for instruction following, coding, and analysis
Flagship Claude model for deep reasoning, coding, and long-horizon agents
xAI's fast agentic tool-calling model with a 2M context window and built-in reasoning
xAI's fast agentic tool-calling model with a 2M context window; non-reasoning variant for low-latency responses
Preview Gemini flagship for complex reasoning, coding, and rich multimodal prompts
Sharper GPT-5 generation for coding, product work, and tool-assisted tasks
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Codex GPT for repository edits, code review, and practical software agents
Kimi reasoning model for long-horizon research, planning, and tool use
Thinking Kimi model for slower research passes, planning, and hard technical questions
Flagship Claude model for deep reasoning, coding, and long-horizon agents
Efficient open MiniMax model built for coding agents and tool-heavy workflows
Fast Claude model for responsive assistance, classification, and lightweight agents
ByteDance Seed model for long-context reasoning, instruction following, and tool-assisted tasks
Fast Claude lane for lightweight agents, office tasks, and responsive chat
Higher-accuracy GPT-5 tier for tough analysis, coding reviews, and planning
Late GLM-4 workhorse for coding agents, reasoning, and structured tasks
Balanced Claude model for coding, analysis, agent workflows, and cost control
Balanced Claude model for coding, analysis, agent workflows, and cost control
Flagship Qwen3 model for coding agents, complex reasoning, and tool use
Qwen vision-language model for visual reasoning, documents, and agent tasks
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Low-latency ByteDance Seed model for high-throughput chat, extraction, and lightweight tool use
Cohere reasoning model for multilingual enterprise agents, tools, and complex workflows
ByteDance Seed multimodal model for image understanding, visual reasoning, and tool-assisted tasks
Original GPT-5 workhorse for reasoning, coding, writing, and tool workflows
Tiny GPT-5 lane for routing, extraction, classification, and bulk jobs
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
Small GPT-5 for responsive agents, coding help, and everyday automation
Flagship Claude model for deep reasoning, coding, and long-horizon agents
Flagship Claude model for deep reasoning, coding, and long-horizon agents
Efficient Qwen model for fast chat, extraction, and high-volume workloads
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| GLM-4.7zhipuai/glm-4.7 | 204.8K | $0.6 | $2.2 | 2025-12-22 | |||
| Gemini 3 Flash Previewgoogle/gemini-3-flash-preview | 1.04858M | $0.5 | $3 | 2025-12-17 | |||
| MiMo-V2-Flashxiaomi/mimo-v2-flash | 262.144K | $0.14 | $0.28 | 2025-12-16 | |||
| Nemotron 3 Nano 30B A3Bnvidia/nemotron-3-nano-30b-a3b | 262.144K | $0.05 | $0.2 | 2025-12-15 | |||
| GPT-5.2 Proopenai/gpt-5.2-pro | 400K | $21 | $168 | 2025-12-11 | |||
| GPT-5.2 Codexopenai/gpt-5.2-codex | 400K | $0.14 | $1.14 | 2025-12-11 | |||
| GPT-5.2openai/gpt-5.2 | 400K | $1.75 | $14 | 2025-12-11 | |||
| Devstral 2mistral/devstral-2512 | 262.144K | $0.4 | $2 | 2025-12-09 | |||
| Devstral Small 2mistral/devstral-small-2 | 262.144K | $0.1 | $0.3 | 2025-12-09 | |||
| Ministral 14Bmistral/ministral-14b | 262.144K | $0.2 | $0.2 | 2025-12-02 | |||
| Mistral Large 3mistral/mistral-large-2512 | 262.144K | $0.5 | $1.5 | 2025-12-02 | |||
| Devstral 2 (latest)mistral/devstral-medium-latest | 262.144K | $0.4 | $2 | 2025-12-02 | |||
| Ministral 3 3Bmistral/ministral-3-3b-instruct-2512 | 262.144K | $0.1 | $0.1 | 2025-12-02 | |||
| Ministral 3 8Bmistral/ministral-3-8b-instruct-2512 | 262.144K | $0.15 | $0.15 | 2025-12-02 | |||
| Ministral 3 14Bmistral/ministral-3-14b-instruct-2512 | 262.144K | $0.1 | $0.4 | 2025-12-02 | |||
| Nova 2 Liteamazon/nova-2-lite | 1M | $0.3 | $2.5 | 2025-12-02 | |||
| DeepSeek Reasonerdeepseek/deepseek-reasoner | 1M | $0.147 | $0.295 | 2025-12-01 | |||
| DeepSeek Chatdeepseek/deepseek-chat | 1M | $0.147 | $0.295 | 2025-12-01 | |||
| Claude Opus 4.5 (latest)anthropic/claude-opus-4-5 | 200K | $5 | $25 | 2025-11-24 | |||
| Grok 4.1 Fast (Reasoning)xai/grok-4.1-fast-reasoning | 2M | $0.2 | $0.5 | 2025-11-19 | |||
| Grok 4.1 Fastxai/grok-4.1-fast | 2M | $0.2 | $0.5 | 2025-11-19 | |||
| Gemini 3 Pro Previewgoogle/gemini-3-pro-preview | 1.04858M | $0.57 | $3.43 | 2025-11-18 | |||
| GPT-5.1openai/gpt-5.1 | 400K | $1.25 | $10 | 2025-11-13 | |||
| GPT-5.1 Codex Maxopenai/gpt-5.1-codex-max | 400K | $1.1 | $9 | 2025-11-13 | |||
| GPT-5.1 Codex miniopenai/gpt-5.1-codex-mini | 400K | $0.22 | $1.8 | 2025-11-13 | |||
| GPT-5.1 Codexopenai/gpt-5.1-codex | 400K | $1.07 | $8.5 | 2025-11-13 | |||
| Kimi K2 Thinking Turbomoonshotai/kimi-k2-thinking-turbo | 262.144K | $1.15 | $8 | 2025-11-06 | |||
| Kimi K2 Thinkingmoonshotai/kimi-k2-thinking | 262.144K | $0.4 | $2.5 | 2025-11-06 | |||
| Claude Opus 4.5anthropic/claude-opus-4-5-20251101 | 200K | $5 | $25 | 2025-11-01 | |||
| MiniMax-M2minimax/MiniMax-M2 | 204.8K | $0.3 | $1.2 | 2025-10-27 | |||
| Claude Haiku 4.5anthropic/claude-haiku-4-5-20251001 | 200K | $1 | $5 | 2025-10-15 | |||
| Seed 1.6bytedance-seed/seed-1-6 | 256K | $0.119 | $1.187 | 2025-10-15 | |||
| Claude Haiku 4.5 (latest)anthropic/claude-haiku-4-5 | 200K | $1 | $5 | 2025-10-15 | |||
| GPT-5 Proopenai/gpt-5-pro | 400K | $15 | $120 | 2025-10-06 | |||
| GLM-4.6zhipuai/glm-4.6 | 204.8K | $0.6 | $2.2 | 2025-09-30 | |||
| Claude Sonnet 4.5 (latest)anthropic/claude-sonnet-4-5 | 200K | $3 | $15 | 2025-09-29 | |||
| Claude Sonnet 4.5anthropic/claude-sonnet-4-5-20250929 | 200K | $3 | $15 | 2025-09-29 | |||
| Qwen3 Maxalibaba/qwen3-max | 262.144K | $1.2 | $6 | 2025-09-23 | |||
| Qwen3-VL Plusalibaba/qwen3-vl-plus | 262.144K | $0.2 | $1.6 | 2025-09-23 | |||
| GPT-5-Codexopenai/gpt-5-codex | 400K | $1.1 | $9 | 2025-09-15 | |||
| Seed 1.6 Flashbytedance-seed/seed-1-6-flash | 256K | $0.022 | $0.223 | 2025-08-28 | |||
| Command A Reasoningcohere/command-a-reasoning-08-2025 | 256K | $2.5 | $10 | 2025-08-21 | |||
| Seed 1.6 Visionbytedance-seed/seed-1-6-vision | 256K | $0.119 | $1.187 | 2025-08-15 | |||
| GPT-5openai/gpt-5 | 400K | $1.25 | $10 | 2025-08-07 | |||
| GPT-5 Nanoopenai/gpt-5-nano | 400K | $0.05 | $0.4 | 2025-08-07 | |||
| GPT-5 Chat (latest)openai/gpt-5-chat-latest | 400K | $1.25 | $10 | 2025-08-07 | |||
| GPT-5 Miniopenai/gpt-5-mini | 400K | $0.25 | $2 | 2025-08-07 | |||
| Claude Opus 4.1 (latest)anthropic/claude-opus-4-1 | 200K | $15 | $75 | 2025-08-05 | |||
| Claude Opus 4.1anthropic/claude-opus-4-1-20250805 | 200K | $15 | $75 | 2025-08-05 | |||
| Qwen Flashalibaba/qwen-flash | 1M | $0.05 | $0.4 | 2025-07-28 |