Strongest Claude Opus model for coding, agents, and professional work
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Frontier GPT-5.6 model for complex professional work, coding, and agentic workflows
Balanced GPT-5.6 model for capable, cost-efficient everyday work
Claude model for creative writing, analysis, and controlled agent workflows
Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
Default frontier GPT for coding, computer use, research, and knowledge work
Everyday Claude agent model for coding, planning, browsing, and general work
Cost-efficient GPT-5.6 model for fast, high-volume workloads
Muse Spark is a natively multimodal reasoning model with support for tool-use, visual chain of thought, and multi-agent orchestration.
Agent-ready GPT for coding and computer-use workflows at a lower cost
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Reasoning-first Gemini preview for agentic coding and complex problem solving
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution
Open MoE flagship with million-token context for coding and long agent runs
Open MiMo model for multimodal coding agents and long-context automation
Strong small GPT for coding subagents, quick tool use, and high-volume work
Cheapest GPT-5.4 lane for simple routing, extraction, and bulk automation
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
Sharper GPT-5 generation for coding, product work, and tool-assisted tasks
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
Largest Gemma 4 instruction model for open, self-hosted chat and reasoning
Nemotron middle tier for collaborative agents and high-volume reasoning workloads
Open GPT reasoning model for self-hosted agents and controllable deployments
Fast NVIDIA Nemotron MoE for reliable agentic tasks across enterprise workloads
Open GPT reasoning model for self-hosted agents and controllable deployments
Affordable GPT-4.1 lane for fast coding help and structured extraction
Small GPT-5 for responsive agents, coding help, and everyday automation
Small Nemotron 3 MoE for efficient coding, math, and long-context agents
Small omni GPT for cheap multimodal assistance and production-scale traffic
Tiny GPT-4.1 option for classification, routing, and very high-volume tasks
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Claude Opus 5anthropic/claude-opus-5 | 78.0 | 1M | $5 | $25 | 2026-07-24 | |||
| GPT-5.6 Solopenai/gpt-5.6-sol | 77.4 | 1.05M | $4 | $20 | 2026-07-09 | |||
| GPT-5.6 Terraopenai/gpt-5.6-terra | 76.7 | 1.05M | $2 | $12 | 2026-07-09 | |||
| Claude Fable 5anthropic/claude-fable-5 | 76.5 | 1M | $10 | $50 | 2026-06-09 | |||
| Kimi K3moonshotai/kimi-k3 | 76.2 | 1.04858M | $3 | $15 | 2026-07-16 | |||
| GPT-5.5openai/gpt-5.5 | 74.9 | 1.05M | $5 | $30 | 2026-04-23 | |||
| Claude Sonnet 5anthropic/claude-sonnet-5 | 71.5 | 1M | $2 | $10 | 2026-06-30 | |||
| GPT-5.6 Lunaopenai/gpt-5.6-luna | 71.4 | 1.05M | $0.2 | $1.2 | 2026-07-09 | |||
| Muse Spark 1.1meta/muse-spark-1.1 | 71.3 | 1.04858M | $1.25 | $4.25 | 2026-04-08 | |||
| GPT-5.4openai/gpt-5.4 | 71.1 | 1.05M | $2.5 | $15 | 2026-03-05 | |||
| Gemini 3.5 Flashgoogle/gemini-3.5-flash | 70.1 | 1.04858M | $1.5 | $9 | 2026-05-19 | |||
| Gemini 3.6 Flashgoogle/gemini-3.6-flash | 69.2 | 1.04858M | $0.75 | $3.75 | 2026-07-21 | |||
| Gemini 3.1 Pro Previewgoogle/gemini-3.1-pro-preview | 68.8 | 1.04858M | $2 | $12 | 2026-02-19 | |||
| Kimi K2.6moonshotai/kimi-k2.6 | 61.8 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| Kimi K2.7 Codemoonshotai/kimi-k2.7-code | 60.8 | 262.144K | $0.95 | $4 | 2026-06-12 | |||
| MiMo-V2.5-Proxiaomi/mimo-v2.5-pro | 60.2 | 1.04858M | $0.435 | $0.87 | 2026-04-22 | |||
| DeepSeek V4 Prodeepseek/deepseek-v4-pro | 59.4 | 1M | $0.435 | $0.87 | 2026-04-24 | |||
| MiMo-V2.5xiaomi/mimo-v2.5 | 56.8 | 1.04858M | $0.14 | $0.28 | 2026-04-22 | |||
| GPT-5.4 miniopenai/gpt-5.4-mini | 56.1 | 400K | $0.75 | $4.5 | 2026-03-17 | |||
| GPT-5.4 nanoopenai/gpt-5.4-nano | 56.1 | 400K | $0.2 | $1.25 | 2026-03-17 | |||
| DeepSeek V4 Flashdeepseek/deepseek-v4-flash | 52.0 | 1M | $0.15 | $0.6 | 2026-04-24 | |||
| GPT-5.1openai/gpt-5.1 | 49.4 | 400K | $1.25 | $10 | 2025-11-13 | |||
| Gemini 3.5 Flash Litegoogle/gemini-3.5-flash-lite | 49.3 | 1.04858M | $0.3 | $2.5 | 2026-07-21 | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 49.3 | 1M | $0.5 | $2.5 | 2026-06-04 | |||
| Gemma 4 31B ITgoogle/gemma-4-31b-it | 43.4 | 262.144K | $0.09 | $0.34 | 2026-04-02 | |||
| Nemotron 3 Super 120B A12Bnvidia/nemotron-3-super-120b-a12b | 37.7 | 262.144K | $0.2 | $0.8 | 2026-03-11 | |||
| GPT OSS 120Bopenai/gpt-oss-120b | 30.4 | 131.072K | $0.03 | $0.17 | 2025-08-05 | |||
| Nemotron 3.5 Lightning 30B A3Bnvidia/nemotron-3.5-lightning | 26.8 | 262.144K | $0.05 | $0.2 | 2026-08-11 | |||
| GPT OSS 20Bopenai/gpt-oss-20b | 20.7 | 131.072K | $0.02 | $0.1 | 2025-08-05 | |||
| GPT-4.1 miniopenai/gpt-4.1-mini | 20.2 | 1.04758M | $0.4 | $1.6 | 2025-04-14 | |||
| GPT-5 Miniopenai/gpt-5-mini | 15.6 | 400K | $0.25 | $2 | 2025-08-07 | |||
| Nemotron 3 Nano 30B A3Bnvidia/nemotron-3-nano-30b-a3b | 14.4 | 262.144K | $0.05 | $0.2 | 2025-12-15 | |||
| GPT-4o miniopenai/gpt-4o-mini | 11.4 | 128K | $0.15 | $0.6 | 2024-07-18 | |||
| GPT-4.1 nanoopenai/gpt-4.1-nano | 11.1 | 1.04758M | $0.1 | $0.4 | 2025-04-14 |