Agent-ready GPT for coding and computer-use workflows at a lower cost
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Strongest Claude Opus model for coding, agents, and professional work
Claude model for creative writing, analysis, and controlled agent workflows
Frontier GPT-5.6 model for complex professional work, coding, and agentic workflows
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
Balanced GPT-5.6 model for capable, cost-efficient everyday work
Default frontier GPT for coding, computer use, research, and knowledge work
Everyday Claude agent model for coding, planning, browsing, and general work
Cost-efficient GPT-5.6 model for fast, high-volume workloads
Sharper GPT-5 generation for coding, product work, and tool-assisted tasks
Muse Spark is a natively multimodal reasoning model with support for tool-use, visual chain of thought, and multi-agent orchestration.
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Open MoE flagship with million-token context for coding and long agent runs
Reasoning-first Gemini preview for agentic coding and complex problem solving
Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
Strong small GPT for coding subagents, quick tool use, and high-volume work
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Open MiMo model for multimodal coding agents and long-context automation
Cheapest GPT-5.4 lane for simple routing, extraction, and bulk automation
Small GPT-5 for responsive agents, coding help, and everyday automation
Largest Gemma 4 instruction model for open, self-hosted chat and reasoning
Affordable GPT-4.1 lane for fast coding help and structured extraction
Nemotron middle tier for collaborative agents and high-volume reasoning workloads
Fast NVIDIA Nemotron MoE for reliable agentic tasks across enterprise workloads
Open GPT reasoning model for self-hosted agents and controllable deployments
Tiny GPT-4.1 option for classification, routing, and very high-volume tasks
Open GPT reasoning model for self-hosted agents and controllable deployments
Small Nemotron 3 MoE for efficient coding, math, and long-context agents
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| GPT-5.4openai/gpt-5.4 | 53.1 | 1.05M | $2.5 | $15 | 2026-03-05 | |||
| Claude Opus 5anthropic/claude-opus-5 | 50.7 | 1M | $5 | $25 | 2026-07-24 | |||
| Claude Fable 5anthropic/claude-fable-5 | 49.7 | 1M | $10 | $50 | 2026-06-09 | |||
| GPT-5.6 Solopenai/gpt-5.6-sol | 47.1 | 1.05M | $4 | $20 | 2026-07-09 | |||
| Kimi K2.6moonshotai/kimi-k2.6 | 45.1 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| Kimi K3moonshotai/kimi-k3 | 43.8 | 1.04858M | $3 | $15 | 2026-07-16 | |||
| GPT-5.6 Terraopenai/gpt-5.6-terra | 42.3 | 1.05M | $2 | $12 | 2026-07-09 | |||
| GPT-5.5openai/gpt-5.5 | 38.6 | 1.05M | $5 | $30 | 2026-04-23 | |||
| Claude Sonnet 5anthropic/claude-sonnet-5 | 38.4 | 1M | $2 | $10 | 2026-06-30 | |||
| GPT-5.6 Lunaopenai/gpt-5.6-luna | 37.5 | 1.05M | $0.2 | $1.2 | 2026-07-09 | |||
| GPT-5.1openai/gpt-5.1 | 37.5 | 400K | $1.25 | $10 | 2025-11-13 | |||
| Muse Spark 1.1meta/muse-spark-1.1 | 34.3 | 1.04858M | $1.25 | $4.25 | 2026-04-08 | |||
| Gemini 3.6 Flashgoogle/gemini-3.6-flash | 34.3 | 1.04858M | $0.75 | $3.75 | 2026-07-21 | |||
| Gemini 3.5 Flashgoogle/gemini-3.5-flash | 33.0 | 1.04858M | $1.5 | $9 | 2026-05-19 | |||
| DeepSeek V4 Prodeepseek/deepseek-v4-pro | 30.9 | 1M | $0.435 | $0.87 | 2026-04-24 | |||
| Gemini 3.1 Pro Previewgoogle/gemini-3.1-pro-preview | 30.4 | 1.04858M | $2 | $12 | 2026-02-19 | |||
| MiMo-V2.5-Proxiaomi/mimo-v2.5-pro | 26.4 | 1.04858M | $0.435 | $0.87 | 2026-04-22 | |||
| Kimi K2.7 Codemoonshotai/kimi-k2.7-code | 26.3 | 262.144K | $0.95 | $4 | 2026-06-12 | |||
| DeepSeek V4 Flashdeepseek/deepseek-v4-flash | 24.8 | 1M | $0.15 | $0.6 | 2026-04-24 | |||
| GPT-5.4 miniopenai/gpt-5.4-mini | 24.6 | 400K | $0.75 | $4.5 | 2026-03-17 | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 23.4 | 1M | $0.5 | $2.5 | 2026-06-04 | |||
| Gemini 3.5 Flash Litegoogle/gemini-3.5-flash-lite | 22.7 | 1.04858M | $0.3 | $2.5 | 2026-07-21 | |||
| MiMo-V2.5xiaomi/mimo-v2.5 | 22.3 | 1.04858M | $0.14 | $0.28 | 2026-04-22 | |||
| GPT-5.4 nanoopenai/gpt-5.4-nano | 21.2 | 400K | $0.2 | $1.25 | 2026-03-17 | |||
| GPT-5 Miniopenai/gpt-5-mini | 17.4 | 400K | $0.25 | $2 | 2025-08-07 | |||
| Gemma 4 31B ITgoogle/gemma-4-31b-it | 15.4 | 262.144K | $0.09 | $0.34 | 2026-04-02 | |||
| GPT-4.1 miniopenai/gpt-4.1-mini | 14.8 | 1.04758M | $0.4 | $1.6 | 2025-04-14 | |||
| Nemotron 3 Super 120B A12Bnvidia/nemotron-3-super-120b-a12b | 13.6 | 262.144K | $0.2 | $0.8 | 2026-03-11 | |||
| Nemotron 3.5 Lightning 30B A3Bnvidia/nemotron-3.5-lightning | 13.6 | 262.144K | $0.05 | $0.2 | 2026-08-11 | |||
| GPT OSS 120Bopenai/gpt-oss-120b | 12.3 | 131.072K | $0.03 | $0.17 | 2025-08-05 | |||
| GPT-4.1 nanoopenai/gpt-4.1-nano | 9.6 | 1.04758M | $0.1 | $0.4 | 2025-04-14 | |||
| GPT OSS 20Bopenai/gpt-oss-20b | 9.0 | 131.072K | $0.02 | $0.1 | 2025-08-05 | |||
| Nemotron 3 Nano 30B A3Bnvidia/nemotron-3-nano-30b-a3b | 8.9 | 262.144K | $0.05 | $0.2 | 2025-12-15 |