Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
63 providers
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
Newer StepFun flash model for faster agents, coding, and multimodal prompts
Largest Gemma 4 instruction model for open, self-hosted chat and reasoning
Open multimodal Gemma instruction model for multilingual text generation and image understanding
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Kimi K3moonshotai/kimi-k3 | 50.6 | 1.04858M | $3 | $15 | 2026-07-16 | |||
| Inklingthinkingmachines/inkling | 24.3 | 1.04858M | $1.87 | $4.68 | 2026-07-15 | |||
| Kimi K2.6moonshotai/kimi-k2.6 | 22.1 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| Step 3.7 Flashstepfun/step-3.7-flash | 21.7 | 256K | $0.185 | $1.11 | 2026-05-29 | |||
| Gemma 4 31B ITgoogle/gemma-4-31b-it | 6.7 | 262.144K | $0.09 | $0.34 | 2026-04-02 | |||
| Gemma 3 12B ITgoogle/gemma-3-12b-it | 0.1 | 131.072K | $0.05 | $0.1 | 2025-03-12 |