Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Strongest Claude Opus model for coding, agents, and professional work
Muse Spark 1.2 is a coding-focused update to Muse Spark 1.1 with improvements in code generation, complex debugging, codebase understanding, and end-to-end developer workflows.
High-efficiency Gemini model for agentic workflows, coding, and multimodal reasoning
Muse Spark 1.3 is a multimodal reasoning model from Meta for long-running agentic, multi-agent, and coding workflows. It improves long-horizon agent collaboration, instruction following, and coding efficiency relative to Muse Spark 1.2.
Claude model for creative writing, analysis, and controlled agent workflows
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Muse Spark is a natively multimodal reasoning model with support for tool-use, visual chain of thought, and multi-agent orchestration.
Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution
Default frontier GPT for coding, computer use, research, and knowledge work
Open MiMo model for multimodal coding agents and long-context automation
Google's most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows
Everyday Claude agent model for coding, planning, browsing, and general work
Reasoning-first Gemini preview for agentic coding and complex problem solving
Agent-ready GPT for coding and computer-use workflows at a lower cost
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Google's proven reasoning model for coding, math, and multimodal analysis
Open MoE flagship with million-token context for coding and long agent runs
Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding
Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
Fast Gemini workhorse for multimodal apps where latency and price matter
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
Long-lived GPT workhorse for coding, instruction following, and production apps
DeepSeek chat model for instruction following, coding, and analysis
Low-latency Gemini model for high-volume multimodal and agent workloads
Affordable GPT-4.1 lane for fast coding help and structured extraction
Tiny GPT-4.1 option for classification, routing, and very high-volume tasks
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Kimi K3moonshotai/kimi-k3 | 1365.0 | 1.04858M | $3 | $15 | 2026-07-16 | |||
| Claude Opus 5anthropic/claude-opus-5 | 1355.0 | 1M | $5 | $25 | 2026-07-24 | |||
| Muse Spark 1.2meta/muse-spark-1.2 | 1354.0 | 1.04858M | $1.25 | $4.25 | 2026-08-05 | |||
| Gemini 3.7 Flashgoogle/gemini-3.7-flash | 1328.0 | 1.04858M | $0.75 | $3.75 | 2026-08-13 | |||
| Muse Spark 1.3meta/muse-spark-1.3 | 1327.0 | 1.04858M | $1.25 | $4.25 | 2026-09-02 | |||
| Claude Fable 5anthropic/claude-fable-5 | 1325.0 | 1M | $10 | $50 | 2026-06-09 | |||
| Gemini 3.6 Flashgoogle/gemini-3.6-flash | 1312.0 | 1.04858M | $0.75 | $3.75 | 2026-07-21 | |||
| Muse Spark 1.1meta/muse-spark-1.1 | 1291.0 | 1.04858M | $1.25 | $4.25 | 2026-04-08 | |||
| MiMo-V2.5-Proxiaomi/mimo-v2.5-pro | 1281.0 | 1.04858M | $0.435 | $0.87 | 2026-04-22 | |||
| GPT-5.5openai/gpt-5.5 | 1275.0 | 1.05M | $5 | $30 | 2026-04-23 | |||
| MiMo-V2.5xiaomi/mimo-v2.5 | 1268.0 | 1.04858M | $0.14 | $0.28 | 2026-04-22 | |||
| Gemini 3.8 Flashgoogle/gemini-3.8-flash | 1262.0 | 1.04858M | $0.75 | $3.75 | 2026-09-02 | |||
| Claude Sonnet 5anthropic/claude-sonnet-5 | 1260.0 | 1M | $2 | $10 | 2026-06-30 | |||
| Gemini 3.1 Pro Previewgoogle/gemini-3.1-pro-preview | 1254.0 | 1.04858M | $2 | $12 | 2026-02-19 | |||
| GPT-5.4openai/gpt-5.4 | 1250.0 | 1.05M | $2.5 | $15 | 2026-03-05 | |||
| Gemini 3.5 Flashgoogle/gemini-3.5-flash | 1250.0 | 1.04858M | $1.5 | $9 | 2026-05-19 | |||
| Gemini 2.5 Progoogle/gemini-2.5-pro | 1237.0 | 1.04858M | $1.25 | $10 | 2025-06-17 | |||
| DeepSeek V4 Prodeepseek/deepseek-v4-pro | 1218.0 | 1M | $0.435 | $0.87 | 2026-04-24 | |||
| DeepSeek V4 Flash 0731deepseek/deepseek-v4-flash-0731 | 1197.0 | 1M | $0.035 | $0.07 | 2026-07-31 | |||
| Inklingthinkingmachines/inkling | 1193.0 | 1.04858M | $1.87 | $4.68 | 2026-07-15 | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 1156.0 | 1M | $0.5 | $2.5 | 2026-06-04 | |||
| Gemini 2.5 Flashgoogle/gemini-2.5-flash | 1149.0 | 1.04858M | $0.3 | $2.5 | 2025-06-17 | |||
| DeepSeek V4 Flashdeepseek/deepseek-v4-flash | 1142.0 | 1M | $0.15 | $0.6 | 2026-04-24 | |||
| GPT-4.1openai/gpt-4.1 | 1118.0 | 1.04758M | $2 | $8 | 2025-04-14 | |||
| DeepSeek Chatdeepseek/deepseek-chat | 1104.0 | 1M | $0.147 | $0.295 | 2025-12-01 | |||
| Gemini 3.1 Flash Lite Previewgoogle/gemini-3.1-flash-lite-preview | 1061.0 | 1.04858M | $0.25 | $1.5 | 2026-03-03 | |||
| GPT-4.1 miniopenai/gpt-4.1-mini | 1048.0 | 1.04758M | $0.4 | $1.6 | 2025-04-14 | |||
| GPT-4.1 nanoopenai/gpt-4.1-nano | 905.0 | 1.04758M | $0.1 | $0.4 | 2025-04-14 |