Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
DeepSeek V4 Pro snapshot with million-token context and support for thinking and non-thinking modes
Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding
Open MoE flagship with million-token context for coding and long agent runs
Multimodal MoE reasoning model (276B total, 12B active) for text, image, and audio
Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
Earlier Kimi frontier model for long-context agents, coding, and multimodal work
Open GPT reasoning model for self-hosted agents and controllable deployments
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Kimi K3moonshotai/kimi-k3 | 50.6 | 1.04858M | $3 | $15 | 2026-07-16 | |||
| DeepSeek V4 Pro 0813deepseek/deepseek-v4-pro-0813 | 42.3 | 1M | $0.442 | $0.884 | 2026-08-12 | |||
| DeepSeek V4 Flash 0731deepseek/deepseek-v4-flash-0731 | 41.7 | 1M | $0.035 | $0.07 | 2026-07-31 | |||
| DeepSeek V4 Prodeepseek/deepseek-v4-pro | 27.7 | 1M | $0.435 | $0.87 | 2026-04-24 | |||
| Inkling Smallthinkingmachines/inkling-small | 25.0 | 1.04858M | $0.45 | $1.2 | 2026-07-30 | |||
| Inklingthinkingmachines/inkling | 24.3 | 1.04858M | $1.87 | $4.68 | 2026-07-15 | |||
| Kimi K2.7 Codemoonshotai/kimi-k2.7-code | 22.5 | 262.144K | $0.95 | $4 | 2026-06-12 | |||
| Kimi K2.6moonshotai/kimi-k2.6 | 22.1 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| Kimi K2.5moonshotai/kimi-k2.5 | 21.7 | 262.144K | $0.3 | $1.9 | 2026-01 | |||
| GPT OSS 120Bopenai/gpt-oss-120b | 6.2 | 131.072K | $0.03 | $0.17 | 2025-08-05 |