Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding
43 providers
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding
Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
Nemotron middle tier for collaborative agents and high-volume reasoning workloads
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| DeepSeek V4 Flash 0731deepseek/deepseek-v4-flash-0731 | 1M | $0.035 | $0.07 | 2026-07-31 | |||
| Kimi K3moonshotai/kimi-k3 | 1.04858M | $3 | $15 | 2026-07-16 | |||
| Kimi K2.7 Codemoonshotai/kimi-k2.7-code | 262.144K | $0.95 | $4 | 2026-06-12 | |||
| Nemotron 3 Super 120B A12Bnvidia/nemotron-3-super-120b-a12b | 262.144K | $0.2 | $0.8 | 2026-03-11 |