Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution
Open MoE flagship with million-token context for coding and long agent runs
Open MiMo model for multimodal coding agents and long-context automation
Multimodal MoE reasoning model (276B total, 12B active) for text, image, and audio
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
Largest Gemma 4 instruction model for open, self-hosted chat and reasoning
Nemotron middle tier for collaborative agents and high-volume reasoning workloads
Open GPT reasoning model for self-hosted agents and controllable deployments
Fast NVIDIA Nemotron MoE for reliable agentic tasks across enterprise workloads
Open GPT reasoning model for self-hosted agents and controllable deployments
Small Nemotron 3 MoE for efficient coding, math, and long-context agents
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Kimi K3moonshotai/kimi-k3 | 76.2 | 1.04858M | $3 | $15 | 2026-07-16 | |||
| Kimi K2.6moonshotai/kimi-k2.6 | 61.8 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| Kimi K2.7 Codemoonshotai/kimi-k2.7-code | 60.8 | 262.144K | $0.95 | $4 | 2026-06-12 | |||
| MiMo-V2.5-Proxiaomi/mimo-v2.5-pro | 60.2 | 1.04858M | $0.435 | $0.87 | 2026-04-22 | |||
| DeepSeek V4 Prodeepseek/deepseek-v4-pro | 59.4 | 1M | $0.435 | $0.87 | 2026-04-24 | |||
| MiMo-V2.5xiaomi/mimo-v2.5 | 56.8 | 1.04858M | $0.14 | $0.28 | 2026-04-22 | |||
| Inkling Smallthinkingmachines/inkling-small | 52.9 | 1.04858M | $0.45 | $1.2 | 2026-07-30 | |||
| DeepSeek V4 Flashdeepseek/deepseek-v4-flash | 52.0 | 1M | $0.15 | $0.6 | 2026-04-24 | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 49.3 | 1M | $0.5 | $2.5 | 2026-06-04 | |||
| Gemma 4 31B ITgoogle/gemma-4-31b-it | 43.4 | 262.144K | $0.09 | $0.34 | 2026-04-02 | |||
| Nemotron 3 Super 120B A12Bnvidia/nemotron-3-super-120b-a12b | 37.7 | 262.144K | $0.2 | $0.8 | 2026-03-11 | |||
| GPT OSS 120Bopenai/gpt-oss-120b | 30.4 | 131.072K | $0.03 | $0.17 | 2025-08-05 | |||
| Nemotron 3.5 Lightning 30B A3Bnvidia/nemotron-3.5-lightning | 26.8 | 262.144K | $0.05 | $0.2 | 2026-08-11 | |||
| GPT OSS 20Bopenai/gpt-oss-20b | 20.7 | 131.072K | $0.02 | $0.1 | 2025-08-05 | |||
| Nemotron 3 Nano 30B A3Bnvidia/nemotron-3-nano-30b-a3b | 14.4 | 262.144K | $0.05 | $0.2 | 2025-12-15 |