Fast Gemini model balancing multimodal reasoning, tool use, and cost
31 providers
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio
Low-latency Gemini model for high-volume multimodal and agent workloads
Fast Gemini workhorse for multimodal apps where latency and price matter
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Gemini 3.5 Flashgoogle/gemini-3.5-flash | 1280.0 | 1.04858M | $1.5 | $9 | 2026-05-19 | |||
| Inklingthinkingmachines/inkling | 1134.0 | 1.04858M | $1.87 | $4.68 | 2026-07-15 | |||
| Gemini 3.1 Flash Lite Previewgoogle/gemini-3.1-flash-lite-preview | 1077.0 | 1.04858M | $0.25 | $1.5 | 2026-03-03 | |||
| Gemini 2.5 Flashgoogle/gemini-2.5-flash | 1045.0 | 1.04858M | $0.3 | $2.5 | 2025-06-17 |