Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Strongest Claude Opus model for coding, agents, and professional work
Claude model for creative writing, analysis, and controlled agent workflows
High-efficiency Gemini model for agentic workflows, coding, and multimodal reasoning
Google's most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
Everyday Claude agent model for coding, planning, browsing, and general work
Open MoE flagship with million-token context for coding and long agent runs
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
Reasoning-first Gemini preview for agentic coding and complex problem solving
Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding
Earlier Kimi frontier model for long-context agents, coding, and multimodal work
Default frontier GPT for coding, computer use, research, and knowledge work
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
New Gemini flash lane bringing frontier-style multimodal reasoning to cheaper runs
Tencent Hy reasoning model for coding, instruction following, and agent tasks
Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
Newer StepFun flash model for faster agents, coding, and multimodal prompts
Agent-ready GPT for coding and computer-use workflows at a lower cost
Reliable GPT generation for broad coding, writing, and tool-assisted product work
Sharper GPT-5 generation for coding, product work, and tool-assisted tasks
Original GPT-5 workhorse for reasoning, coding, writing, and tool workflows
Low-latency Gemini model for high-volume multimodal and agent workloads
Small GPT-5 for responsive agents, coding help, and everyday automation
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Tiny GPT-5 lane for routing, extraction, classification, and bulk jobs
Open GPT reasoning model for self-hosted agents and controllable deployments
Fast o-series model for compact reasoning, coding, and tool use
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Kimi K3moonshotai/kimi-k3 | 1426.0 | 1.04858M | $3 | $15 | 2026-07-16 | |||
| Claude Opus 5anthropic/claude-opus-5 | 1364.0 | 1M | $5 | $25 | 2026-07-24 | |||
| Claude Fable 5anthropic/claude-fable-5 | 1347.0 | 1M | $10 | $50 | 2026-06-09 | |||
| Gemini 3.7 Flashgoogle/gemini-3.7-flash | 1336.0 | 1.04858M | $0.75 | $3.75 | 2026-08-13 | |||
| Gemini 3.8 Flashgoogle/gemini-3.8-flash | 1322.0 | 1.04858M | $0.75 | $3.75 | 2026-09-02 | |||
| Gemini 3.6 Flashgoogle/gemini-3.6-flash | 1302.0 | 1.04858M | $0.75 | $3.75 | 2026-07-21 | |||
| Kimi K2.6moonshotai/kimi-k2.6 | 1297.0 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| Claude Sonnet 5anthropic/claude-sonnet-5 | 1288.0 | 1M | $2 | $10 | 2026-06-30 | |||
| DeepSeek V4 Prodeepseek/deepseek-v4-pro | 1283.0 | 1M | $0.435 | $0.87 | 2026-04-24 | |||
| Gemini 3.5 Flashgoogle/gemini-3.5-flash | 1272.0 | 1.04858M | $1.5 | $9 | 2026-05-19 | |||
| Kimi K2.7 Codemoonshotai/kimi-k2.7-code | 1269.0 | 262.144K | $0.95 | $4 | 2026-06-12 | |||
| Gemini 3.1 Pro Previewgoogle/gemini-3.1-pro-preview | 1263.0 | 1.04858M | $2 | $12 | 2026-02-19 | |||
| DeepSeek V4 Flash 0731deepseek/deepseek-v4-flash-0731 | 1238.0 | 1M | $0.035 | $0.07 | 2026-07-31 | |||
| Kimi K2.5moonshotai/kimi-k2.5 | 1234.0 | 262.144K | $0.3 | $1.9 | 2026-01 | |||
| GPT-5.5openai/gpt-5.5 | 1228.0 | 1.05M | $5 | $30 | 2026-04-23 | |||
| DeepSeek V4 Flashdeepseek/deepseek-v4-flash | 1217.0 | 1M | $0.15 | $0.6 | 2026-04-24 | |||
| Gemini 3 Flash Previewgoogle/gemini-3-flash-preview | 1212.0 | 1.04858M | $0.5 | $3 | 2025-12-17 | |||
| Hy3tencent/hy3 | 1205.0 | 256K | $0.066 | $0.26 | 2026-07-06 | |||
| Inklingthinkingmachines/inkling | 1188.0 | 1.04858M | $1.87 | $4.68 | 2026-07-15 | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 1163.0 | 1M | $0.5 | $2.5 | 2026-06-04 | |||
| Step 3.7 Flashstepfun/step-3.7-flash | 1156.0 | 256K | $0.185 | $1.11 | 2026-05-29 | |||
| GPT-5.4openai/gpt-5.4 | 1130.0 | 1.05M | $2.5 | $15 | 2026-03-05 | |||
| GPT-5.2openai/gpt-5.2 | 1108.0 | 400K | $1.75 | $14 | 2025-12-11 | |||
| GPT-5.1openai/gpt-5.1 | 1090.0 | 400K | $1.25 | $10 | 2025-11-13 | |||
| GPT-5openai/gpt-5 | 1084.0 | 400K | $1.25 | $10 | 2025-08-07 | |||
| Gemini 3.1 Flash Lite Previewgoogle/gemini-3.1-flash-lite-preview | 1076.0 | 1.04858M | $0.25 | $1.5 | 2026-03-03 | |||
| GPT-5 Miniopenai/gpt-5-mini | 1066.0 | 400K | $0.25 | $2 | 2025-08-07 | |||
| GPT-5.3 Codexopenai/gpt-5.3-codex | 1037.0 | 400K | $1.75 | $14 | 2026-02-05 | |||
| GPT-5.1 Codex miniopenai/gpt-5.1-codex-mini | 1017.0 | 400K | $0.22 | $1.8 | 2025-11-13 | |||
| GPT-5 Nanoopenai/gpt-5-nano | 994.0 | 400K | $0.05 | $0.4 | 2025-08-07 | |||
| GPT OSS 120Bopenai/gpt-oss-120b | 930.0 | 131.072K | $0.03 | $0.17 | 2025-08-05 | |||
| o4-miniopenai/o4-mini | 882.0 | 200K | $1.1 | $4.4 | 2025-04-16 |