Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution
Open MoE flagship with million-token context for coding and long agent runs
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
Earlier Kimi frontier model for long-context agents, coding, and multimodal work
Largest Gemma 4 instruction model for open, self-hosted chat and reasoning
Open Gemma instruction model for efficient chat and self-hosted deployments
Thinking Kimi model for slower research passes, planning, and hard technical questions
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Kimi K3moonshotai/kimi-k3 | 76.2 | 1.04858M | $3 | $15 | 2026-07-16 | |||
| Kimi K2.6moonshotai/kimi-k2.6 | 61.8 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| Kimi K2.7 Codemoonshotai/kimi-k2.7-code | 60.8 | 262.144K | $0.95 | $4 | 2026-06-12 | |||
| MiMo-V2.5-Proxiaomi/mimo-v2.5-pro | 60.2 | 1.04858M | $0.435 | $0.87 | 2026-04-22 | |||
| DeepSeek V4 Prodeepseek/deepseek-v4-pro | 59.4 | 1M | $0.435 | $0.87 | 2026-04-24 | |||
| DeepSeek V4 Flashdeepseek/deepseek-v4-flash | 52.0 | 1M | $0.15 | $0.6 | 2026-04-24 | |||
| Kimi K2.5moonshotai/kimi-k2.5 | 46.8 | 262.144K | $0.3 | $1.9 | 2026-01 | |||
| Gemma 4 31B ITgoogle/gemma-4-31b-it | 43.4 | 262.144K | $0.09 | $0.34 | 2026-04-02 | |||
| Gemma 4 26B A4B ITgoogle/gemma-4-26b-a4b-it | 39.3 | 262.144K | $0.042 | $0.22 | 2026-04-02 | |||
| Kimi K2 Thinkingmoonshotai/kimi-k2-thinking | 21.0 | 262.144K | $0.4 | $2.5 | 2025-11-06 |