Strongest Claude Opus model for coding, agents, and professional work
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
GPT-6 Astra is OpenAI's most capable model for complex reasoning, coding, computer use, research, and document creation.
Claude model for creative writing, analysis, and controlled agent workflows
Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
Frontier GPT-5.6 model for complex professional work, coding, and agentic workflows
Everyday Claude agent model for coding, planning, browsing, and general work
Agent-ready GPT for coding and computer-use workflows at a lower cost
Muse Spark 1.2 is a coding-focused update to Muse Spark 1.1 with improvements in code generation, complex debugging, codebase understanding, and end-to-end developer workflows.
Balanced GPT-5.6 model for capable, cost-efficient everyday work
Cost-efficient GPT-5.6 model for fast, high-volume workloads
DeepSeek V4 Pro snapshot with million-token context and support for thinking and non-thinking modes
Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding
Google's most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows
Default frontier GPT for coding, computer use, research, and knowledge work
High-efficiency Gemini model for agentic workflows, coding, and multimodal reasoning
Upstage's flagship model, specialized for agentic use
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
Open MoE flagship with million-token context for coding and long agent runs
Muse Spark is a natively multimodal reasoning model with support for tool-use, visual chain of thought, and multi-agent orchestration.
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Original GPT-5 workhorse for reasoning, coding, writing, and tool workflows
Tencent Hy reasoning model for coding, instruction following, and agent tasks
Multimodal MoE reasoning model (276B total, 12B active) for text, image, and audio
Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio
Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
Newer StepFun flash model for faster agents, coding, and multimodal prompts
Earlier Kimi frontier model for long-context agents, coding, and multimodal work
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
Sharper GPT-5 generation for coding, product work, and tool-assisted tasks
Strong small GPT for coding subagents, quick tool use, and high-volume work
Hybrid-reasoning DeepSeek model with thinking and non-thinking modes, sparse attention, and tool-use
Cheapest GPT-5.4 lane for simple routing, extraction, and bulk automation
Open MiMo model for multimodal coding agents and long-context automation
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Meituan LongCat-2.0, a reasoning model with tool calling and a 1M-token context window
Open Gemma instruction model for efficient chat and self-hosted deployments
Reasoning-first Gemini preview for agentic coding and complex problem solving
Small GPT-5 for responsive agents, coding help, and everyday automation
Largest Gemma 4 instruction model for open, self-hosted chat and reasoning
Open GPT reasoning model for self-hosted agents and controllable deployments
Fast NVIDIA Nemotron MoE for reliable agentic tasks across enterprise workloads
Nemotron middle tier for collaborative agents and high-volume reasoning workloads
Google's proven reasoning model for coding, math, and multimodal analysis
Low-latency Gemini model for high-volume multimodal and agent workloads
Thinking Kimi model for slower research passes, planning, and hard technical questions
Open GPT reasoning model for self-hosted agents and controllable deployments
Reasoning-optimized 398B MoE agent model with extended thinking for long-horizon and multi-turn tool use
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Claude Opus 5anthropic/claude-opus-5 | 56.2 | 1M | $5 | $25 | 2026-07-24 | |||
| GPT-6 Astraopenai/gpt-6-astra | 51.5 | 1.05M | $10 | $50 | 2026-09-04 | |||
| Claude Fable 5anthropic/claude-fable-5 | 51.0 | 1M | $10 | $50 | 2026-06-09 | |||
| Kimi K3moonshotai/kimi-k3 | 50.6 | 1.04858M | $3 | $15 | 2026-07-16 | |||
| GPT-5.6 Solopenai/gpt-5.6-sol | 50.5 | 1.05M | $4 | $20 | 2026-07-09 | |||
| Claude Sonnet 5anthropic/claude-sonnet-5 | 44.3 | 1M | $2 | $10 | 2026-06-30 | |||
| GPT-5.4openai/gpt-5.4 | 44.2 | 1.05M | $2.5 | $15 | 2026-03-05 | |||
| Muse Spark 1.2meta/muse-spark-1.2 | 44.0 | 1.04858M | $1.25 | $4.25 | 2026-08-05 | |||
| GPT-5.6 Terraopenai/gpt-5.6-terra | 43.7 | 1.05M | $2 | $12 | 2026-07-09 | |||
| GPT-5.6 Lunaopenai/gpt-5.6-luna | 42.7 | 1.05M | $0.2 | $1.2 | 2026-07-09 | |||
| DeepSeek V4 Pro 0813deepseek/deepseek-v4-pro-0813 | 42.3 | 1M | $0.442 | $0.884 | 2026-08-12 | |||
| DeepSeek V4 Flash 0731deepseek/deepseek-v4-flash-0731 | 41.7 | 1M | $0.035 | $0.07 | 2026-07-31 | |||
| Gemini 3.8 Flashgoogle/gemini-3.8-flash | 41.1 | 1.04858M | $0.75 | $3.75 | 2026-09-02 | |||
| GPT-5.5openai/gpt-5.5 | 37.3 | 1.05M | $5 | $30 | 2026-04-23 | |||
| Gemini 3.7 Flashgoogle/gemini-3.7-flash | 36.4 | 1.04858M | $0.75 | $3.75 | 2026-08-13 | |||
| Solar Pro 4upstage/solar-pro4 | 33.6 | 524.288K | $0.3 | $1.2 | 2026-08-06 | |||
| Gemini 3.6 Flashgoogle/gemini-3.6-flash | 30.2 | 1.04858M | $0.75 | $3.75 | 2026-07-21 | |||
| DeepSeek V4 Flashdeepseek/deepseek-v4-flash | 27.9 | 1M | $0.15 | $0.6 | 2026-04-24 | |||
| DeepSeek V4 Prodeepseek/deepseek-v4-pro | 27.7 | 1M | $0.435 | $0.87 | 2026-04-24 | |||
| Muse Spark 1.1meta/muse-spark-1.1 | 27.5 | 1.04858M | $1.25 | $4.25 | 2026-04-08 | |||
| Gemini 3.5 Flashgoogle/gemini-3.5-flash | 27.3 | 1.04858M | $1.5 | $9 | 2026-05-19 | |||
| GPT-5openai/gpt-5 | 26.5 | 400K | $1.25 | $10 | 2025-08-07 | |||
| Hy3 previewtencent/hy3-preview | 25.6 | 256K | $0.066 | $0.26 | 2026-04-20 | |||
| Inkling Smallthinkingmachines/inkling-small | 25.0 | 1.04858M | $0.45 | $1.2 | 2026-07-30 | |||
| Inklingthinkingmachines/inkling | 24.3 | 1.04858M | $1.87 | $4.68 | 2026-07-15 | |||
| MiMo-V2.5-Proxiaomi/mimo-v2.5-pro | 22.7 | 1.04858M | $0.435 | $0.87 | 2026-04-22 | |||
| Kimi K2.7 Codemoonshotai/kimi-k2.7-code | 22.5 | 262.144K | $0.95 | $4 | 2026-06-12 | |||
| Kimi K2.6moonshotai/kimi-k2.6 | 22.1 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| Step 3.7 Flashstepfun/step-3.7-flash | 21.7 | 256K | $0.185 | $1.11 | 2026-05-29 | |||
| Kimi K2.5moonshotai/kimi-k2.5 | 21.7 | 262.144K | $0.3 | $1.9 | 2026-01 | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 21.7 | 1M | $0.5 | $2.5 | 2026-06-04 | |||
| GPT-5.1openai/gpt-5.1 | 21.6 | 400K | $1.25 | $10 | 2025-11-13 | |||
| GPT-5.4 miniopenai/gpt-5.4-mini | 19.7 | 400K | $0.75 | $4.5 | 2026-03-17 | |||
| DeepSeek V3.2deepseek/deepseek-v3.2 | 18.3 | 128K | $0.18 | $0.35 | 2025-12-01 | |||
| GPT-5.4 nanoopenai/gpt-5.4-nano | 17.7 | 400K | $0.2 | $1.25 | 2026-03-17 | |||
| MiMo-V2.5xiaomi/mimo-v2.5 | 17.4 | 1.04858M | $0.14 | $0.28 | 2026-04-22 | |||
| Gemini 3.5 Flash Litegoogle/gemini-3.5-flash-lite | 15.9 | 1.04858M | $0.3 | $2.5 | 2026-07-21 | |||
| LongCat-2.0meituan/longcat-2.0 | 15.9 | 1M | $0.3 | $1.2 | 2026-06-30 | |||
| Gemma 4 26B A4B ITgoogle/gemma-4-26b-a4b-it | 11.0 | 262.144K | $0.042 | $0.22 | 2026-04-02 | |||
| Gemini 3.1 Pro Previewgoogle/gemini-3.1-pro-preview | 10.3 | 1.04858M | $2 | $12 | 2026-02-19 | |||
| GPT-5 Miniopenai/gpt-5-mini | 8.9 | 400K | $0.25 | $2 | 2025-08-07 | |||
| Gemma 4 31B ITgoogle/gemma-4-31b-it | 6.7 | 262.144K | $0.09 | $0.34 | 2026-04-02 | |||
| GPT OSS 120Bopenai/gpt-oss-120b | 6.2 | 131.072K | $0.03 | $0.17 | 2025-08-05 | |||
| Nemotron 3.5 Lightning 30B A3Bnvidia/nemotron-3.5-lightning | 6.1 | 262.144K | $0.05 | $0.2 | 2026-08-11 | |||
| Nemotron 3 Super 120B A12Bnvidia/nemotron-3-super-120b-a12b | 4.1 | 262.144K | $0.2 | $0.8 | 2026-03-11 | |||
| Gemini 2.5 Progoogle/gemini-2.5-pro | 3.5 | 1.04858M | $1.25 | $10 | 2025-06-17 | |||
| Gemini 3.1 Flash Lite Previewgoogle/gemini-3.1-flash-lite-preview | 3.2 | 1.04858M | $0.25 | $1.5 | 2026-03-03 | |||
| Kimi K2 Thinkingmoonshotai/kimi-k2-thinking | 1.8 | 262.144K | $0.4 | $2.5 | 2025-11-06 | |||
| GPT OSS 20Bopenai/gpt-oss-20b | 1.4 | 131.072K | $0.02 | $0.1 | 2025-08-05 | |||
| Trinity Large Thinkingarcee-ai/trinity-large-thinking | 1.2 | 524.288K | $0.25 | $0.9 | 2026-04-01 |