DeepSeek V4 Pro snapshot with million-token context and support for thinking and non-thinking modes
Model catalog
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
30 providers
Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding
43 providers
Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
63 providers
Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio
21 providers
Open flagship GLM for long-horizon coding agents and million-token context work
72 providers
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
15 providers
MiniMax multimodal model for long-context coding, perception, and agent planning
37 providers
Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks
31 providers
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| DeepSeek V4 Pro 0813deepseek/deepseek-v4-pro-0813 | 1M | $0.442 | $0.884 | 2026-08-12 | |||
| DeepSeek V4 Flash 0731deepseek/deepseek-v4-flash-0731 | 1M | $0.035 | $0.07 | 2026-07-31 | |||
| Kimi K3moonshotai/kimi-k3 | 1.04858M | $3 | $15 | 2026-07-16 | |||
| Inklingthinkingmachines/inkling | 1.04858M | $1.87 | $4.68 | 2026-07-15 | |||
| GLM-5.2zhipuai/glm-5.2 | 1M | $1.4 | $4.4 | 2026-06-13 | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 1M | $0.5 | $2.5 | 2026-06-04 | |||
| MiniMax-M3minimax/MiniMax-M3 | 1.04858M | $0.3 | $1.2 | 2026-06-01 | |||
| Qwen3.7 Maxalibaba/qwen3.7-max | 1M | $2.5 | $7.5 | 2026-05-21 |