Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Strongest Claude Opus model for coding, agents, and professional work
High-efficiency Gemini model for agentic workflows, coding, and multimodal reasoning
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
Google's most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows
Everyday Claude agent model for coding, planning, browsing, and general work
Agent-ready GPT for coding and computer-use workflows at a lower cost
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
Google's proven reasoning model for coding, math, and multimodal analysis
Original GPT-5 workhorse for reasoning, coding, writing, and tool workflows
Earlier Kimi frontier model for long-context agents, coding, and multimodal work
Sharper GPT-5 generation for coding, product work, and tool-assisted tasks
Open MoE flagship with million-token context for coding and long agent runs
Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding
Hybrid-reasoning DeepSeek model with thinking and non-thinking modes, sparse attention, and tool-use
Fast Gemini workhorse for multimodal apps where latency and price matter
Small GPT-5 for responsive agents, coding help, and everyday automation
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
Tencent Hy reasoning model for coding, instruction following, and agent tasks
Long-lived GPT workhorse for coding, instruction following, and production apps
Tiny GPT-5 lane for routing, extraction, classification, and bulk jobs
Affordable GPT-4.1 lane for fast coding help and structured extraction
Open GPT reasoning model for self-hosted agents and controllable deployments
Open GPT reasoning model for self-hosted agents and controllable deployments
Tiny GPT-4.1 option for classification, routing, and very high-volume tasks
Omni-era GPT for multimodal chat, practical coding, and general assistants
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Kimi K3moonshotai/kimi-k3 | 1365.0 | 1.04858M | $3 | $15 | 2026-07-16 | |||
| Claude Opus 5anthropic/claude-opus-5 | 1355.0 | 1M | $5 | $25 | 2026-07-24 | |||
| Gemini 3.7 Flashgoogle/gemini-3.7-flash | 1328.0 | 1.04858M | $0.75 | $3.75 | 2026-08-13 | |||
| Gemini 3.6 Flashgoogle/gemini-3.6-flash | 1312.0 | 1.04858M | $0.75 | $3.75 | 2026-07-21 | |||
| Kimi K2.6moonshotai/kimi-k2.6 | 1266.0 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| Gemini 3.8 Flashgoogle/gemini-3.8-flash | 1262.0 | 1.04858M | $0.75 | $3.75 | 2026-09-02 | |||
| Claude Sonnet 5anthropic/claude-sonnet-5 | 1260.0 | 1M | $2 | $10 | 2026-06-30 | |||
| GPT-5.4openai/gpt-5.4 | 1250.0 | 1.05M | $2.5 | $15 | 2026-03-05 | |||
| Gemini 3.5 Flashgoogle/gemini-3.5-flash | 1250.0 | 1.04858M | $1.5 | $9 | 2026-05-19 | |||
| Kimi K2.7 Codemoonshotai/kimi-k2.7-code | 1238.0 | 262.144K | $0.95 | $4 | 2026-06-12 | |||
| Gemini 2.5 Progoogle/gemini-2.5-pro | 1237.0 | 1.04858M | $1.25 | $10 | 2025-06-17 | |||
| GPT-5openai/gpt-5 | 1237.0 | 400K | $1.25 | $10 | 2025-08-07 | |||
| Kimi K2.5moonshotai/kimi-k2.5 | 1235.0 | 262.144K | $0.3 | $1.9 | 2026-01 | |||
| GPT-5.1openai/gpt-5.1 | 1220.0 | 400K | $1.25 | $10 | 2025-11-13 | |||
| DeepSeek V4 Prodeepseek/deepseek-v4-pro | 1218.0 | 1M | $0.435 | $0.87 | 2026-04-24 | |||
| DeepSeek V4 Flash 0731deepseek/deepseek-v4-flash-0731 | 1197.0 | 1M | $0.05 | $0.16 | 2026-07-31 | |||
| DeepSeek V3.2deepseek/deepseek-v3.2 | 1174.0 | 128K | $0.18 | $0.35 | 2025-12-01 | |||
| Gemini 2.5 Flashgoogle/gemini-2.5-flash | 1149.0 | 1.04858M | $0.3 | $2.5 | 2025-06-17 | |||
| GPT-5 Miniopenai/gpt-5-mini | 1146.0 | 400K | $0.25 | $2 | 2025-08-07 | |||
| DeepSeek V4 Flashdeepseek/deepseek-v4-flash | 1142.0 | 1M | $0.15 | $0.6 | 2026-04-24 | |||
| Hy3tencent/hy3 | 1141.0 | 256K | $0.066 | $0.26 | 2026-07-06 | |||
| GPT-4.1openai/gpt-4.1 | 1118.0 | 1.04758M | $2 | $8 | 2025-04-14 | |||
| GPT-5 Nanoopenai/gpt-5-nano | 1071.0 | 400K | $0.05 | $0.4 | 2025-08-07 | |||
| GPT-4.1 miniopenai/gpt-4.1-mini | 1048.0 | 1.04758M | $0.4 | $1.6 | 2025-04-14 | |||
| GPT OSS 120Bopenai/gpt-oss-120b | 1002.0 | 131.072K | $0.03 | $0.17 | 2025-08-05 | |||
| GPT OSS 20Bopenai/gpt-oss-20b | 947.0 | 131.072K | $0.02 | $0.1 | 2025-08-05 | |||
| GPT-4.1 nanoopenai/gpt-4.1-nano | 905.0 | 1.04758M | $0.1 | $0.4 | 2025-04-14 | |||
| GPT-4oopenai/gpt-4o | 871.0 | 128K | $2.5 | $10 | 2024-05-13 |