Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Strongest Claude Opus model for coding, agents, and professional work
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Claude model for creative writing, analysis, and controlled agent workflows
Everyday Claude agent model for coding, planning, browsing, and general work
Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
Muse Spark is a natively multimodal reasoning model with support for tool-use, visual chain of thought, and multi-agent orchestration.
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
Open MiMo model for multimodal coding agents and long-context automation
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Default frontier GPT for coding, computer use, research, and knowledge work
Reasoning-first Gemini preview for agentic coding and complex problem solving
Open MoE flagship with million-token context for coding and long agent runs
Agent-ready GPT for coding and computer-use workflows at a lower cost
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
New Gemini flash lane bringing frontier-style multimodal reasoning to cheaper runs
Sharper GPT-5 generation for coding, product work, and tool-assisted tasks
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
Small GPT-5 for responsive agents, coding help, and everyday automation
Tiny GPT-5 lane for routing, extraction, classification, and bulk jobs
Long-lived GPT workhorse for coding, instruction following, and production apps
Affordable GPT-4.1 lane for fast coding help and structured extraction
Tiny GPT-4.1 option for classification, routing, and very high-volume tasks
Open GPT reasoning model for self-hosted agents and controllable deployments
Open GPT reasoning model for self-hosted agents and controllable deployments
Omni-era GPT for multimodal chat, practical coding, and general assistants
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Kimi K3moonshotai/kimi-k3 | 1354.0 | 1.04858M | $3 | $15 | 2026-07-16 | |||
| Claude Opus 5anthropic/claude-opus-5 | 1320.0 | 1M | $5 | $25 | 2026-07-24 | |||
| Gemini 3.6 Flashgoogle/gemini-3.6-flash | 1312.0 | 1.04858M | $0.75 | $3.75 | 2026-07-21 | |||
| Claude Fable 5anthropic/claude-fable-5 | 1310.0 | 1M | $10 | $50 | 2026-06-09 | |||
| Claude Sonnet 5anthropic/claude-sonnet-5 | 1289.0 | 1M | $2 | $10 | 2026-06-30 | |||
| MiMo-V2.5-Proxiaomi/mimo-v2.5-pro | 1285.0 | 1.04858M | $0.435 | $0.87 | 2026-04-22 | |||
| Kimi K2.6moonshotai/kimi-k2.6 | 1282.0 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| Muse Spark 1.1meta/muse-spark-1.1 | 1282.0 | 1.04858M | $1.25 | $4.25 | 2026-04-08 | |||
| Kimi K2.7 Codemoonshotai/kimi-k2.7-code | 1279.0 | 262.144K | $0.95 | $4 | 2026-06-12 | |||
| MiMo-V2.5xiaomi/mimo-v2.5 | 1279.0 | 1.04858M | $0.14 | $0.28 | 2026-04-22 | |||
| Gemini 3.5 Flashgoogle/gemini-3.5-flash | 1275.0 | 1.04858M | $1.5 | $9 | 2026-05-19 | |||
| GPT-5.5openai/gpt-5.5 | 1269.0 | 1.05M | $5 | $30 | 2026-04-23 | |||
| Gemini 3.1 Pro Previewgoogle/gemini-3.1-pro-preview | 1265.0 | 1.04858M | $2 | $12 | 2026-02-19 | |||
| DeepSeek V4 Prodeepseek/deepseek-v4-pro | 1247.0 | 1M | $0.435 | $0.87 | 2026-04-24 | |||
| GPT-5.4openai/gpt-5.4 | 1232.0 | 1.05M | $2.5 | $15 | 2026-03-05 | |||
| DeepSeek V4 Flashdeepseek/deepseek-v4-flash | 1220.0 | 1M | $0.15 | $0.6 | 2026-04-24 | |||
| Gemini 3 Flash Previewgoogle/gemini-3-flash-preview | 1208.0 | 1.04858M | $0.5 | $3 | 2025-12-17 | |||
| GPT-5.1openai/gpt-5.1 | 1199.0 | 400K | $1.25 | $10 | 2025-11-13 | |||
| GPT-5.3 Codexopenai/gpt-5.3-codex | 1175.0 | 400K | $1.75 | $14 | 2026-02-05 | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 1144.0 | 1M | $0.5 | $2.5 | 2026-06-04 | |||
| GPT-5 Miniopenai/gpt-5-mini | 1137.0 | 400K | $0.25 | $2 | 2025-08-07 | |||
| GPT-5 Nanoopenai/gpt-5-nano | 1114.0 | 400K | $0.05 | $0.4 | 2025-08-07 | |||
| GPT-4.1openai/gpt-4.1 | 1051.0 | 1.04758M | $2 | $8 | 2025-04-14 | |||
| GPT-4.1 miniopenai/gpt-4.1-mini | 1010.0 | 1.04758M | $0.4 | $1.6 | 2025-04-14 | |||
| GPT-4.1 nanoopenai/gpt-4.1-nano | 985.0 | 1.04758M | $0.1 | $0.4 | 2025-04-14 | |||
| GPT OSS 120Bopenai/gpt-oss-120b | 980.0 | 131.072K | $0.03 | $0.17 | 2025-08-05 | |||
| GPT OSS 20Bopenai/gpt-oss-20b | 865.0 | 131.072K | $0.02 | $0.1 | 2025-08-05 | |||
| GPT-4oopenai/gpt-4o | 843.0 | 128K | $2.5 | $10 | 2024-05-13 |