Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
Model catalog
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
63 providers
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
64 providers
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
61 providers
Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding
43 providers
Earlier Kimi frontier model for long-context agents, coding, and multimodal work
46 providers
Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio
21 providers
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
15 providers
Open GPT reasoning model for self-hosted agents and controllable deployments
53 providers
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Kimi K3moonshotai/kimi-k3 | 1426.0 | 1.04858M | $3 | $15 | 2026-07-16 | |||
| Kimi K2.6moonshotai/kimi-k2.6 | 1297.0 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| Kimi K2.7 Codemoonshotai/kimi-k2.7-code | 1269.0 | 262.144K | $0.95 | $4 | 2026-06-12 | |||
| DeepSeek V4 Flash 0731deepseek/deepseek-v4-flash-0731 | 1238.0 | 1M | $0.035 | $0.07 | 2026-07-31 | |||
| Kimi K2.5moonshotai/kimi-k2.5 | 1234.0 | 262.144K | $0.3 | $1.9 | 2026-01 | |||
| Inklingthinkingmachines/inkling | 1188.0 | 1.04858M | $1.87 | $4.68 | 2026-07-15 | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 1163.0 | 1M | $0.5 | $2.5 | 2026-06-04 | |||
| GPT OSS 120Bopenai/gpt-oss-120b | 930.0 | 131.072K | $0.03 | $0.17 | 2025-08-05 |