Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
Model catalog
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
63 providers
Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio
21 providers
MiniMax multimodal model for long-context coding, perception, and agent planning
37 providers
Newer StepFun flash model for faster agents, coding, and multimodal prompts
17 providers
Balanced Mistral model for enterprise assistants, multilingual work, and tools
9 providers
Open Nemotron omni model combining reasoning with text, vision, and audio
5 providers
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
64 providers
Largest Gemma 4 instruction model for open, self-hosted chat and reasoning
33 providers
Qwen vision-language model for visual reasoning, documents, and agent tasks
16 providers
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Kimi K3moonshotai/kimi-k3 | 1.04858M | $3 | $15 | 2026-07-16 | |||
| Inklingthinkingmachines/inkling | 1.04858M | $1.87 | $4.68 | 2026-07-15 | |||
| MiniMax-M3minimax/MiniMax-M3 | 1.04858M | $0.3 | $1.2 | 2026-06-01 | |||
| Step 3.7 Flashstepfun/step-3.7-flash | 256K | $0.185 | $1.11 | 2026-05-29 | |||
| Mistral Medium 3.5mistral/mistral-medium-2604 | 262.144K | $1.5 | $7.5 | 2026-04-29 | |||
| Nemotron 3 Nano Omni 30B A3B Reasoningnvidia/nemotron-3-nano-omni-30b-a3b-reasoning | 256K | $0.2 | $0.8 | 2026-04-28 | |||
| Kimi K2.6moonshotai/kimi-k2.6 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| Gemma 4 31B ITgoogle/gemma-4-31b-it | 262.144K | $0.09 | $0.34 | 2026-04-02 | |||
| Qwen3.5 122B-A10Balibaba/qwen3.5-122b-a10b | 262.144K | $0.4 | $3.2 | 2026-02-23 |