Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding
43 providers
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding
Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
Open flagship GLM for long-horizon coding agents and million-token context work
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
MiniMax multimodal model for long-context coding, perception, and agent planning
Open MoE flagship with million-token context for coding and long agent runs
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| DeepSeek V4 Flash 0731deepseek/deepseek-v4-flash-0731 | 1M | $0.035 | $0.07 | 2026-07-31 | |||
| Kimi K3moonshotai/kimi-k3 | 1.04858M | $3 | $15 | 2026-07-16 | |||
| GLM-5.2zhipuai/glm-5.2 | 1M | $1.4 | $4.4 | 2026-06-13 | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 1M | $0.5 | $2.5 | 2026-06-04 | |||
| MiniMax-M3minimax/MiniMax-M3 | 1.04858M | $0.3 | $1.2 | 2026-06-01 | |||
| DeepSeek V4 Prodeepseek/deepseek-v4-pro | 1M | $0.435 | $0.87 | 2026-04-24 | |||
| DeepSeek V4 Flashdeepseek/deepseek-v4-flash | 1M | $0.15 | $0.6 | 2026-04-24 |