Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
63 providers
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding
Open MoE flagship with million-token context for coding and long agent runs
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Kimi K3moonshotai/kimi-k3 | 1352.0 | 1.04858M | $3 | $15 | 2026-07-16 | |||
| DeepSeek V4 Flash 0731deepseek/deepseek-v4-flash-0731 | 1250.0 | 1M | $0.035 | $0.07 | 2026-07-31 | |||
| DeepSeek V4 Prodeepseek/deepseek-v4-pro | 1247.0 | 1M | $0.435 | $0.87 | 2026-04-24 | |||
| DeepSeek V4 Flashdeepseek/deepseek-v4-flash | 1220.0 | 1M | $0.15 | $0.6 | 2026-04-24 | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 1145.0 | 1M | $0.5 | $2.5 | 2026-06-04 |