DeepSeek V4.1 Flash model for reasoning and agentic coding
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Google's most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows
Claude model for demanding reasoning and long-horizon agentic work
High-efficiency Gemini model for agentic workflows, coding, and multimodal reasoning
DeepSeek V4 Pro snapshot with million-token context and support for thinking and non-thinking modes
xAI's frontier model for long-running agents, coding, knowledge work, and visual projects
Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding
Strongest Claude Opus model for coding, agents, and professional work
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio
Frontier GPT-5.6 model for complex professional work, coding, and agentic workflows
Balanced GPT-5.6 model for capable, cost-efficient everyday work
Cost-efficient GPT-5.6 model for fast, high-volume workloads
xAI's Grok model for chat, coding, agentic tools, and lower hallucination risk
Everyday Claude agent model for coding, planning, browsing, and general work
Faster ByteDance Seed 2.1 model for multimodal reasoning and latency-sensitive agent workflows
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
Claude model for creative writing, analysis, and controlled agent workflows
Top Claude Opus tier for the hardest reasoning, coding, and long-horizon agents
Default frontier GPT for coding, computer use, research, and knowledge work
Open MiMo model for multimodal coding agents and long-context automation
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
xAI's default Grok for chat, coding, agentic tools, and lower hallucination risk
Stronger Opus tier for advanced software work and high-stakes reasoning
Fast Grok coding model tuned for agentic engineering and iterative edits
Nemotron model for efficient reasoning, coding, and specialized AI agents
Fast Mistral production model for chat, extraction, and cost-sensitive agents
Agent-ready GPT for coding and computer-use workflows at a lower cost
Claude workhorse for coding agents, careful analysis, and production cost control
High-end Claude for difficult coding, planning, and slower expert reasoning
Earlier Kimi frontier model for long-context agents, coding, and multimodal work
Reliable GPT generation for broad coding, writing, and tool-assisted product work
Hybrid-reasoning DeepSeek model with thinking and non-thinking modes, sparse attention, and tool-use
Flagship Claude model for deep reasoning, coding, and long-horizon agents
Balanced Claude model for coding, analysis, agent workflows, and cost control
Omni-era GPT for multimodal chat, practical coding, and general assistants
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| DeepSeek V4.1 Flashdeepseek/deepseek-v4.1-flash | 1M | $0.15 | $0.6 | 2026-09-10 | |||
| Gemini 3.8 Flashgoogle/gemini-3.8-flash | 1.04858M | $0.75 | $3.75 | 2026-09-02 | |||
| Claude Fable 5.1anthropic/claude-fable-5-1 | 1M | $10 | $50 | 2026-09-01 | |||
| Gemini 3.7 Flashgoogle/gemini-3.7-flash | 1.04858M | $0.75 | $3.75 | 2026-08-13 | |||
| DeepSeek V4 Pro 0813deepseek/deepseek-v4-pro-0813 | 1M | $0.442 | $0.884 | 2026-08-12 | |||
| Grok 4.6xai/grok-4.6 | 500K | $2 | $6 | 2026-08-12 | |||
| DeepSeek V4 Flash 0731deepseek/deepseek-v4-flash-0731 | 1M | $0.035 | $0.07 | 2026-07-31 | |||
| Claude Opus 5anthropic/claude-opus-5 | 1M | $5 | $25 | 2026-07-24 | |||
| Gemini 3.6 Flashgoogle/gemini-3.6-flash | 1.04858M | $0.75 | $3.75 | 2026-07-21 | |||
| Kimi K3moonshotai/kimi-k3 | 1.04858M | $3 | $15 | 2026-07-16 | |||
| Inklingthinkingmachines/inkling | 1.04858M | $1.87 | $4.68 | 2026-07-15 | |||
| GPT-5.6 Solopenai/gpt-5.6-sol | 1.05M | $4 | $20 | 2026-07-09 | |||
| GPT-5.6 Terraopenai/gpt-5.6-terra | 1.05M | $2 | $12 | 2026-07-09 | |||
| GPT-5.6 Lunaopenai/gpt-5.6-luna | 1.05M | $0.2 | $1.2 | 2026-07-09 | |||
| Grok 4.5xai/grok-4.5 | 500K | $2 | $6 | 2026-07-08 | |||
| Claude Sonnet 5anthropic/claude-sonnet-5 | 1M | $2 | $10 | 2026-06-30 | |||
| Seed 2.1 Turbobytedance-seed/seed-2.1-turbo | 256K | $0.354 | $1.77 | 2026-06-23 | |||
| Kimi K2.7 Codemoonshotai/kimi-k2.7-code | 262.144K | $0.95 | $4 | 2026-06-12 | |||
| Claude Fable 5anthropic/claude-fable-5 | 1M | $10 | $50 | 2026-06-09 | |||
| Claude Opus 4.8anthropic/claude-opus-4-8 | 1M | $5 | $25 | 2026-05-28 | |||
| GPT-5.5openai/gpt-5.5 | 1.05M | $5 | $30 | 2026-04-23 | |||
| MiMo-V2.5xiaomi/mimo-v2.5 | 1.04858M | $0.14 | $0.28 | 2026-04-22 | |||
| Kimi K2.6moonshotai/kimi-k2.6 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| Grok 4.3xai/grok-4.3 | 1M | $1.25 | $2.5 | 2026-04-17 | |||
| Claude Opus 4.7anthropic/claude-opus-4-7 | 1M | $5 | $25 | 2026-04-16 | |||
| Grok Build 0.1xai/grok-build-0.1 | 256K | $1 | $2 | 2026-04-16 | |||
| Nemotron Cascade 2 30B A3Bnvidia/nemotron-cascade-2-30b-a3b | 256K | — | — | 2026-03-24 | |||
| Mistral Small 4mistral/mistral-small-2603 | 256K | $0.15 | $0.6 | 2026-03-16 | |||
| GPT-5.4openai/gpt-5.4 | 1.05M | $2.5 | $15 | 2026-03-05 | |||
| Claude Sonnet 4.6anthropic/claude-sonnet-4-6 | 1M | $3 | $15 | 2026-02-17 | |||
| Claude Opus 4.6anthropic/claude-opus-4-6 | 1M | $5 | $25 | 2026-02-05 | |||
| Kimi K2.5moonshotai/kimi-k2.5 | 262.144K | $0.3 | $1.9 | 2026-01 | |||
| GPT-5.2openai/gpt-5.2 | 400K | $1.75 | $14 | 2025-12-11 | |||
| DeepSeek V3.2deepseek/deepseek-v3.2 | 128K | $0.18 | $0.35 | 2025-12-01 | |||
| Claude Opus 4.5anthropic/claude-opus-4-5-20251101 | 200K | $5 | $25 | 2025-11-01 | |||
| Claude Sonnet 4.5anthropic/claude-sonnet-4-5-20250929 | 200K | $3 | $15 | 2025-09-29 | |||
| GPT-4oopenai/gpt-4o | 128K | $2.5 | $10 | 2024-05-13 |