Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Open MoE flagship with million-token context for coding and long agent runs
Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution
Open MiMo model for multimodal coding agents and long-context automation
Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
DeepSeek chat model for instruction following, coding, and analysis
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Kimi K3moonshotai/kimi-k3 | 1426.0 | 1.04858M | $3 | $15 | 2026-07-16 | |||
| DeepSeek V4 Prodeepseek/deepseek-v4-pro | 1283.0 | 1M | $0.435 | $0.87 | 2026-04-24 | |||
| MiMo-V2.5-Proxiaomi/mimo-v2.5-pro | 1276.0 | 1.04858M | $0.435 | $0.87 | 2026-04-22 | |||
| MiMo-V2.5xiaomi/mimo-v2.5 | 1241.0 | 1.04858M | $0.14 | $0.28 | 2026-04-22 | |||
| DeepSeek V4 Flash 0731deepseek/deepseek-v4-flash-0731 | 1238.0 | 1M | $0.05 | $0.16 | 2026-07-31 | |||
| DeepSeek V4 Flashdeepseek/deepseek-v4-flash | 1217.0 | 1M | $0.15 | $0.6 | 2026-04-24 | |||
| Inklingthinkingmachines/inkling | 1188.0 | 1.04858M | $1.87 | $4.68 | 2026-07-15 | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 1163.0 | 1M | $0.5 | $2.5 | 2026-06-04 | |||
| DeepSeek Chatdeepseek/deepseek-chat | 1115.0 | 1M | $0.147 | $0.295 | 2025-12-01 |