Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
Open MoE flagship with million-token context for coding and long agent runs
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding
Earlier Kimi frontier model for long-context agents, coding, and multimodal work
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio
Hybrid-reasoning DeepSeek model with thinking and non-thinking modes, sparse attention, and tool-use
Open GPT reasoning model for self-hosted agents and controllable deployments
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Kimi K3moonshotai/kimi-k3 | 1426.0 | 1.04858M | $3 | $15 | 2026-07-16 | |||
| Kimi K2.6moonshotai/kimi-k2.6 | 1297.0 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| DeepSeek V4 Prodeepseek/deepseek-v4-pro | 1283.0 | 1M | $0.435 | $0.87 | 2026-04-24 | |||
| Kimi K2.7 Codemoonshotai/kimi-k2.7-code | 1269.0 | 262.144K | $0.95 | $4 | 2026-06-12 | |||
| DeepSeek V4 Flash 0731deepseek/deepseek-v4-flash-0731 | 1238.0 | 1M | $0.05 | $0.16 | 2026-07-31 | |||
| Kimi K2.5moonshotai/kimi-k2.5 | 1234.0 | 262.144K | $0.3 | $1.9 | 2026-01 | |||
| DeepSeek V4 Flashdeepseek/deepseek-v4-flash | 1217.0 | 1M | $0.15 | $0.6 | 2026-04-24 | |||
| Inklingthinkingmachines/inkling | 1188.0 | 1.04858M | $1.87 | $4.68 | 2026-07-15 | |||
| DeepSeek V3.2deepseek/deepseek-v3.2 | 1161.0 | 128K | $0.18 | $0.35 | 2025-12-01 | |||
| GPT OSS 120Bopenai/gpt-oss-120b | 930.0 | 131.072K | $0.03 | $0.17 | 2025-08-05 |