Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
Model catalog
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
63 providers
Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding
43 providers
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
62 providers
Open MoE flagship with million-token context for coding and long agent runs
62 providers
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
61 providers
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
15 providers
Fast NVIDIA Nemotron MoE for reliable agentic tasks across enterprise workloads
10 providers
Nemotron middle tier for collaborative agents and high-volume reasoning workloads
15 providers
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Kimi K3moonshotai/kimi-k3 | 50.6 | 1.04858M | $3 | $15 | 2026-07-16 | |||
| DeepSeek V4 Flash 0731deepseek/deepseek-v4-flash-0731 | 41.7 | 1M | $0.035 | $0.07 | 2026-07-31 | |||
| DeepSeek V4 Flashdeepseek/deepseek-v4-flash | 27.9 | 1M | $0.15 | $0.6 | 2026-04-24 | |||
| DeepSeek V4 Prodeepseek/deepseek-v4-pro | 27.7 | 1M | $0.435 | $0.87 | 2026-04-24 | |||
| Kimi K2.7 Codemoonshotai/kimi-k2.7-code | 22.5 | 262.144K | $0.95 | $4 | 2026-06-12 | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 21.7 | 1M | $0.5 | $2.5 | 2026-06-04 | |||
| Nemotron 3.5 Lightning 30B A3Bnvidia/nemotron-3.5-lightning | 6.1 | 262.144K | $0.05 | $0.2 | 2026-08-11 | |||
| Nemotron 3 Super 120B A12Bnvidia/nemotron-3-super-120b-a12b | 4.1 | 262.144K | $0.2 | $0.8 | 2026-03-11 |