Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution
Open MiMo model for multimodal coding agents and long-context automation
Open MoE flagship with million-token context for coding and long agent runs
Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding
Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
DeepSeek chat model for instruction following, coding, and analysis
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Kimi K3moonshotai/kimi-k3 | 1365.0 | 1.04858M | $3 | $15 | 2026-07-16 | |||
| MiMo-V2.5-Proxiaomi/mimo-v2.5-pro | 1281.0 | 1.04858M | $0.435 | $0.87 | 2026-04-22 | |||
| MiMo-V2.5xiaomi/mimo-v2.5 | 1268.0 | 1.04858M | $0.14 | $0.28 | 2026-04-22 | |||
| DeepSeek V4 Prodeepseek/deepseek-v4-pro | 1218.0 | 1M | $0.435 | $0.87 | 2026-04-24 | |||
| DeepSeek V4 Flash 0731deepseek/deepseek-v4-flash-0731 | 1197.0 | 1M | $0.035 | $0.07 | 2026-07-31 | |||
| Inklingthinkingmachines/inkling | 1193.0 | 1.04858M | $1.87 | $4.68 | 2026-07-15 | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 1156.0 | 1M | $0.5 | $2.5 | 2026-06-04 | |||
| DeepSeek V4 Flashdeepseek/deepseek-v4-flash | 1142.0 | 1M | $0.15 | $0.6 | 2026-04-24 | |||
| DeepSeek Chatdeepseek/deepseek-chat | 1104.0 | 1M | $0.147 | $0.295 | 2025-12-01 |