Strongest Claude Opus model for coding, agents, and professional work
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Frontier GPT-5.6 model for complex professional work, coding, and agentic workflows
GPT-6 Astra is OpenAI's most capable model for complex reasoning, coding, computer use, research, and document creation.
Balanced GPT-5.6 model for capable, cost-efficient everyday work
Claude model for creative writing, analysis, and controlled agent workflows
Google's most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows
Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
High-efficiency Gemini model for agentic workflows, coding, and multimodal reasoning
Default frontier GPT for coding, computer use, research, and knowledge work
Everyday Claude agent model for coding, planning, browsing, and general work
Cost-efficient GPT-5.6 model for fast, high-volume workloads
Agent-ready GPT for coding and computer-use workflows at a lower cost
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding
Reasoning-first Gemini preview for agentic coding and complex problem solving
DeepSeek V4 Pro snapshot with million-token context and support for thinking and non-thinking modes
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
Open MoE flagship with million-token context for coding and long agent runs
Cheapest GPT-5.4 lane for simple routing, extraction, and bulk automation
Strong small GPT for coding subagents, quick tool use, and high-volume work
Multimodal MoE reasoning model (276B total, 12B active) for text, image, and audio
Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
Sharper GPT-5 generation for coding, product work, and tool-assisted tasks
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
Earlier Kimi frontier model for long-context agents, coding, and multimodal work
O-series reasoning model for hard analysis, math, coding, and planning
Newer StepFun flash model for faster agents, coding, and multimodal prompts
Original GPT-5 workhorse for reasoning, coding, writing, and tool workflows
Nemotron middle tier for collaborative agents and high-volume reasoning workloads
Low-latency Gemini model for high-volume multimodal and agent workloads
Open GPT reasoning model for self-hosted agents and controllable deployments
Classic open reasoning model for transparent math, coding, and deliberate problem solving
Compact GPT model for low-latency assistance and high-volume workloads
Thinking Kimi model for slower research passes, planning, and hard technical questions
Open GPT reasoning model for self-hosted agents and controllable deployments
Affordable GPT-4.1 lane for fast coding help and structured extraction
Small GPT-5 for responsive agents, coding help, and everyday automation
Small Nemotron 3 MoE for efficient coding, math, and long-context agents
GPT model for general reasoning, writing, coding, and tool-assisted tasks
Small omni GPT for cheap multimodal assistance and production-scale traffic
Tiny GPT-4.1 option for classification, routing, and very high-volume tasks
Compact GPT model for low-latency assistance and high-volume workloads
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Claude Opus 5anthropic/claude-opus-5 | 78.0 | 1M | $5 | $25 | 2026-07-24 | |||
| GPT-5.6 Solopenai/gpt-5.6-sol | 77.4 | 1.05M | $4 | $20 | 2026-07-09 | |||
| GPT-6 Astraopenai/gpt-6-astra | 76.9 | 1.05M | $10 | $50 | 2026-09-04 | |||
| GPT-5.6 Terraopenai/gpt-5.6-terra | 76.7 | 1.05M | $2 | $12 | 2026-07-09 | |||
| Claude Fable 5anthropic/claude-fable-5 | 76.5 | 1M | $10 | $50 | 2026-06-09 | |||
| Gemini 3.8 Flashgoogle/gemini-3.8-flash | 76.3 | 1.04858M | $0.75 | $3.75 | 2026-09-02 | |||
| Kimi K3moonshotai/kimi-k3 | 76.2 | 1.04858M | $3 | $15 | 2026-07-16 | |||
| Gemini 3.7 Flashgoogle/gemini-3.7-flash | 76.1 | 1.04858M | $0.75 | $3.75 | 2026-08-13 | |||
| GPT-5.5openai/gpt-5.5 | 74.9 | 1.05M | $5 | $30 | 2026-04-23 | |||
| Claude Sonnet 5anthropic/claude-sonnet-5 | 71.5 | 1M | $2 | $10 | 2026-06-30 | |||
| GPT-5.6 Lunaopenai/gpt-5.6-luna | 71.4 | 1.05M | $0.2 | $1.2 | 2026-07-09 | |||
| GPT-5.4openai/gpt-5.4 | 71.1 | 1.05M | $2.5 | $15 | 2026-03-05 | |||
| Gemini 3.5 Flashgoogle/gemini-3.5-flash | 70.1 | 1.04858M | $1.5 | $9 | 2026-05-19 | |||
| Gemini 3.6 Flashgoogle/gemini-3.6-flash | 69.2 | 1.04858M | $0.75 | $3.75 | 2026-07-21 | |||
| DeepSeek V4 Flash 0731deepseek/deepseek-v4-flash-0731 | 69.1 | 1M | $0.05 | $0.16 | 2026-07-31 | |||
| Gemini 3.1 Pro Previewgoogle/gemini-3.1-pro-preview | 68.8 | 1.04858M | $2 | $12 | 2026-02-19 | |||
| DeepSeek V4 Pro 0813deepseek/deepseek-v4-pro-0813 | 68.8 | 1M | $0.442 | $0.884 | 2026-08-12 | |||
| Kimi K2.6moonshotai/kimi-k2.6 | 61.8 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| Kimi K2.7 Codemoonshotai/kimi-k2.7-code | 60.8 | 262.144K | $0.95 | $4 | 2026-06-12 | |||
| DeepSeek V4 Prodeepseek/deepseek-v4-pro | 59.4 | 1M | $0.435 | $0.87 | 2026-04-24 | |||
| GPT-5.4 nanoopenai/gpt-5.4-nano | 56.1 | 400K | $0.2 | $1.25 | 2026-03-17 | |||
| GPT-5.4 miniopenai/gpt-5.4-mini | 56.1 | 400K | $0.75 | $4.5 | 2026-03-17 | |||
| Inkling Smallthinkingmachines/inkling-small | 52.9 | 1.04858M | $0.45 | $1.2 | 2026-07-30 | |||
| Inklingthinkingmachines/inkling | 52.1 | 1.04858M | $1.87 | $4.68 | 2026-07-15 | |||
| DeepSeek V4 Flashdeepseek/deepseek-v4-flash | 52.0 | 1M | $0.15 | $0.6 | 2026-04-24 | |||
| GPT-5.1openai/gpt-5.1 | 49.4 | 400K | $1.25 | $10 | 2025-11-13 | |||
| Gemini 3.5 Flash Litegoogle/gemini-3.5-flash-lite | 49.3 | 1.04858M | $0.3 | $2.5 | 2026-07-21 | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 49.3 | 1M | $0.5 | $2.5 | 2026-06-04 | |||
| Kimi K2.5moonshotai/kimi-k2.5 | 46.8 | 262.144K | $0.3 | $1.9 | 2026-01 | |||
| o1openai/o1 | 39.7 | 200K | $15 | $60 | 2024-12-05 | |||
| Step 3.7 Flashstepfun/step-3.7-flash | 39.6 | 256K | $0.185 | $1.11 | 2026-05-29 | |||
| GPT-5openai/gpt-5 | 37.8 | 400K | $1.25 | $10 | 2025-08-07 | |||
| Nemotron 3 Super 120B A12Bnvidia/nemotron-3-super-120b-a12b | 37.7 | 262.144K | $0.2 | $0.8 | 2026-03-11 | |||
| Gemini 3.1 Flash Lite Previewgoogle/gemini-3.1-flash-lite-preview | 34.7 | 1.04858M | $0.25 | $1.5 | 2026-03-03 | |||
| GPT OSS 120Bopenai/gpt-oss-120b | 30.4 | 131.072K | $0.03 | $0.17 | 2025-08-05 | |||
| DeepSeek-R1deepseek/deepseek-r1 | 24.6 | 128K | $0.7 | $2.5 | 2025-01-20 | |||
| GPT-4 Turboopenai/gpt-4-turbo | 21.5 | 128K | $10 | $30 | 2023-11-06 | |||
| Kimi K2 Thinkingmoonshotai/kimi-k2-thinking | 21.0 | 262.144K | $0.4 | $2.5 | 2025-11-06 | |||
| GPT OSS 20Bopenai/gpt-oss-20b | 20.7 | 131.072K | $0.02 | $0.1 | 2025-08-05 | |||
| GPT-4.1 miniopenai/gpt-4.1-mini | 20.2 | 1.04758M | $0.4 | $1.6 | 2025-04-14 | |||
| GPT-5 Miniopenai/gpt-5-mini | 15.6 | 400K | $0.25 | $2 | 2025-08-07 | |||
| Nemotron 3 Nano 30B A3Bnvidia/nemotron-3-nano-30b-a3b | 14.4 | 262.144K | $0.05 | $0.2 | 2025-12-15 | |||
| GPT-4openai/gpt-4 | 13.1 | 8.192K | $30 | $60 | 2023-11-06 | |||
| GPT-4o miniopenai/gpt-4o-mini | 11.4 | 128K | $0.15 | $0.6 | 2024-07-18 | |||
| GPT-4.1 nanoopenai/gpt-4.1-nano | 11.1 | 1.04758M | $0.1 | $0.4 | 2025-04-14 | |||
| GPT-3.5-turboopenai/gpt-3.5-turbo | 10.7 | 16.385K | $0.5 | $1.5 | 2023-03-01 |