Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
64 providers
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
Open MoE flagship with million-token context for coding and long agent runs
Newer StepFun flash model for faster agents, coding, and multimodal prompts
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio
Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Kimi K2.6moonshotai/kimi-k2.6 | 1183.0 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| DeepSeek V4 Prodeepseek/deepseek-v4-pro | 1167.0 | 1M | $0.435 | $0.87 | 2026-04-24 | |||
| Step 3.7 Flashstepfun/step-3.7-flash | 1165.0 | 256K | $0.185 | $1.11 | 2026-05-29 | |||
| DeepSeek V4 Flashdeepseek/deepseek-v4-flash | 1128.0 | 1M | $0.15 | $0.6 | 2026-04-24 | |||
| Inklingthinkingmachines/inkling | 1118.0 | 1.04858M | $1.87 | $4.68 | 2026-07-15 | |||
| DeepSeek V4 Flash 0731deepseek/deepseek-v4-flash-0731 | 1110.0 | 1M | $0.035 | $0.07 | 2026-07-31 | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 1099.0 | 1M | $0.5 | $2.5 | 2026-06-04 |