Multimodal MoE reasoning model (276B total, 12B active) for text, image, and audio
11 providers
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Multimodal MoE reasoning model (276B total, 12B active) for text, image, and audio
Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio
Open MiMo model for multimodal coding agents and long-context automation
Open multimodal Qwen MoE for local agents that need vision, audio, and code
Large open Qwen multimodal MoE for visual agents and long technical tasks
Instruct model with native audio input for speech understanding and tool use
Lean Gemini 2.5 lane for cheap multimodal traffic and quick agents
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Inkling Smallthinkingmachines/inkling-small | 1.04858M | $0.45 | $1.2 | 2026-07-30 | |||
| Inklingthinkingmachines/inkling | 1.04858M | $1.87 | $4.68 | 2026-07-15 | |||
| MiMo-V2.5xiaomi/mimo-v2.5 | 1.04858M | $0.14 | $0.28 | 2026-04-22 | |||
| Qwen3.6 35B-A3Balibaba/qwen3.6-35b-a3b | 262.144K | $0.248 | $1.485 | 2026-04-17 | |||
| Qwen3.5 397B-A17Balibaba/qwen3.5-397b-a17b | 262.144K | $0.6 | $3.6 | 2026-02-15 | |||
| Voxtral Small (latest)mistral/voxtral-small-latest | 32K | $0.1 | $0.3 | 2025-07-15 | |||
| Gemini 2.5 Flash-Litegoogle/gemini-2.5-flash-lite | 1.04858M | $0.1 | $0.4 | 2025-06-17 |