Google's most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Muse Spark 1.3 is a multimodal reasoning model from Meta for long-running agentic, multi-agent, and coding workflows. It improves long-horizon agent collaboration, instruction following, and coding efficiency relative to Muse Spark 1.2.
Speech transcription model for accurate audio-to-text and captioning workflows
High-efficiency Gemini model for agentic workflows, coding, and multimodal reasoning
High-efficiency Gemini model for agentic workflows, coding, and multimodal reasoning
Muse Spark 1.2 is a coding-focused update to Muse Spark 1.1 with improvements in code generation, complex debugging, codebase understanding, and end-to-end developer workflows.
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Realtime speech-to-speech model with configurable reasoning, tool use, and robust voice-agent behavior
Low-latency audio-to-audio model for real-time speech translation across 70+ languages
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Low-latency Gemini model for high-volume multimodal and agent workloads
Streaming speech-to-text model for low-latency transcript deltas from live audio
Multimodal embedding model mapping text, images, video, audio, and PDFs into a unified embedding space
Maximum-comprehensiveness agentic researcher for multi-step investigation, synthesis, and cited reports
Agentic model for autonomous multi-step research, synthesis, and cited reports
Low-latency speech generation with steerable prompts and expressive audio tags
Vision-language model for embodied reasoning: spatial understanding, task planning, and physical-world agentic robotics
High-quality, low-latency Live API model for real-time dialogue and voice-first AI applications
Music generation model for short 30-second clips, loops, and previews from text or image prompts
Music generation model for full-length songs from text or images with vocals and structure
MiMo omni model for text, image, video, audio, and agents
Low-latency Gemini model for high-volume multimodal and agent workloads
Advanced Gemini model for complex reasoning, coding, and multimodal analysis
Reasoning-first Gemini preview for agentic coding and complex problem solving
New Gemini flash lane bringing frontier-style multimodal reasoning to cheaper runs
Preview Gemini flagship for complex reasoning, coding, and rich multimodal prompts
Speech generation model for controllable voice, narration, and audio delivery
Speech generation model for controllable voice, narration, and audio delivery
Fast Gemini workhorse for multimodal apps where latency and price matter
Lean Gemini 2.5 lane for cheap multimodal traffic and quick agents
Google's proven reasoning model for coding, math, and multimodal analysis
Qwen omni model for text, vision, audio, and multimodal agent tasks
Earlier Gemini Flash workhorse for responsive multimodal apps and tool use
Low-latency Gemini model for high-volume multimodal and agent workloads
Muse Spark 1.2 contributor tier is a reasoning model from Meta designed for developers who want to start building at an even lower cost. It’s meaningfully cheaper than Muse Spark...
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...
The Auto Router automatically selects the best model for your prompt, powered by the wisdom of the market. It routes you based on what the OpenRouter community collectively spends on...
Voxtral Small is an enhancement of Mistral Small 3, incorporating state-of-the-art audio input capabilities while retaining best-in-class text performance. It excels at speech transcription, translation and audio understanding. Input audio...
This model always redirects to the latest model in the Google Gemini Flash family.
Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...
This model always redirects to the latest model in the Google Gemini Pro family.
A cost-efficient version of GPT Audio. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Input is priced at $0.60 per million...
NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
The gpt-audio model is OpenAI's first generally available audio model. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Audio is priced...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Gemini 3.8 Flashgoogle/gemini-3.8-flash | 1.04858M | $0.75 | $3.75 | 2026-09-02 | |||
| Muse Spark 1.3meta/muse-spark-1.3 | 1.04858M | $1.25 | $4.25 | 2026-09-02 | |||
| Gemini 3.5 Transcribe Livegoogle/gemini-3.5-transcribe-live | Not documented | — | — | 2026-08-26 | |||
| Gemini 3.7 Flashgoogle/gemini-3.7-flash | 1.04858M | $0.75 | $3.75 | 2026-08-13 | |||
| Gemini Flash Latestgoogle/gemini-flash-latest | 1.04858M | $0.75 | $3.75 | 2026-08-13 | |||
| Muse Spark 1.2meta/muse-spark-1.2 | 1.04858M | $1.25 | $4.25 | 2026-08-05 | |||
| Gemini Flash-Lite Latestgoogle/gemini-flash-lite-latest | 1.04858M | $0.3 | $2.5 | 2026-07-21 | |||
| Gemini 3.5 Flash Litegoogle/gemini-3.5-flash-lite | 1.04858M | $0.3 | $2.5 | 2026-07-21 | |||
| Gemini 3.6 Flashgoogle/gemini-3.6-flash | 1.04858M | $0.75 | $3.75 | 2026-07-21 | |||
| GPT-Realtime-2.1openai/gpt-realtime-2.1 | 128K | $4 | $24 | 2026-07-06 | |||
| Gemini 3.5 Live Translate Previewgoogle/gemini-3.5-live-translate-preview | 131.072K | $3.5 | $21 | 2026-06-09 | |||
| Gemini 3.5 Flashgoogle/gemini-3.5-flash | 1.04858M | $1.5 | $9 | 2026-05-19 | |||
| Gemini 3.1 Flash Litegoogle/gemini-3.1-flash-lite | 1.04858M | $0.25 | $1.5 | 2026-05-07 | |||
| GPT Realtime Whisperopenai/gpt-realtime-whisper | Not documented | — | — | 2026-05-07 | |||
| Gemini Embedding 2google/gemini-embedding-2 | 8.192K | $0.2 | — | 2026-04-22 | |||
| Deep Research Max Previewgoogle/deep-research-max-preview-04-2026 | 1.04858M | — | — | 2026-04-21 | |||
| Gemini Deep Research Previewgoogle/deep-research-preview-04-2026 | 1.04858M | — | — | 2026-04-21 | |||
| Gemini 3.1 Flash TTS Previewgoogle/gemini-3.1-flash-tts-preview | 8.192K | $1 | $20 | 2026-04-15 | |||
| Gemini Robotics-ER 1.6 Previewgoogle/gemini-robotics-er-1.6-preview | 131.072K | $1 | $5 | 2026-04-14 | |||
| Gemini 3.1 Flash Live Previewgoogle/gemini-3.1-flash-live-preview | 131.072K | $0.75 | $4.5 | 2026-03-26 | |||
| Lyria 3 Clip Previewgoogle/lyria-3-clip-preview | 131.072K | — | — | 2026-03-25 | |||
| Lyria 3 Pro Previewgoogle/lyria-3-pro-preview | 131.072K | — | — | 2026-03-25 | |||
| MiMo-V2-Omnixiaomi/mimo-v2-omni | 262.144K | $0.14 | $0.28 | 2026-03-18 | |||
| Gemini 3.1 Flash Lite Previewgoogle/gemini-3.1-flash-lite-preview | 1.04858M | $0.25 | $1.5 | 2026-03-03 | |||
| Gemini 3.1 Pro Preview Custom Toolsgoogle/gemini-3.1-pro-preview-customtools | 1.04858M | $2 | $12 | 2026-02-19 | |||
| Gemini 3.1 Pro Previewgoogle/gemini-3.1-pro-preview | 1.04858M | $2 | $12 | 2026-02-19 | |||
| Gemini 3 Flash Previewgoogle/gemini-3-flash-preview | 1.04858M | $0.5 | $3 | 2025-12-17 | |||
| Gemini 3 Pro Previewgoogle/gemini-3-pro-preview | 1.04858M | $0.57 | $3.43 | 2025-11-18 | |||
| Gemini 2.5 Flash TTSgoogle/gemini-2.5-flash-tts | 32.768K | $0.5 | $10 | 2025-09-30 | |||
| Gemini 2.5 Pro TTSgoogle/gemini-2.5-pro-tts | 32.768K | $1 | $20 | 2025-09-30 | |||
| Gemini 2.5 Flashgoogle/gemini-2.5-flash | 1.04858M | $0.3 | $2.5 | 2025-06-17 | |||
| Gemini 2.5 Flash-Litegoogle/gemini-2.5-flash-lite | 1.04858M | $0.1 | $0.4 | 2025-06-17 | |||
| Gemini 2.5 Progoogle/gemini-2.5-pro | 1.04858M | $1.25 | $10 | 2025-06-17 | |||
| Qwen-Omni Turboalibaba/qwen-omni-turbo | 32.768K | $0.07 | $0.27 | 2025-01-19 | |||
| Gemini 2.0 Flashgoogle/gemini-2.0-flash | 1.04858M | $0.1 | $0.42 | 2024-12-11 | |||
| Gemini 2.0 Flash-Litegoogle/gemini-2.0-flash-lite | 1.04858M | $0.052 | $0.21 | 2024-12-11 | |||
| Meta: Muse Spark 1.2 Contributormeta/muse-spark-1.2-contributor | 1.04858M | $0.1 | $0.2 | — | |||
| Thinking Machines: Inkling (free)thinkingmachines/inkling:free | 1.04858M | Free | Free | — | |||
| Google: Gemini 3.6 Flash (batch)google/gemini-3.6-flash:batch | 1.04858M | $0.375 | $1.875 | — | |||
| Auto Routeropenrouter/auto | 2M | — | — | — | |||
| Mistral: Voxtral Small 24B 2507mistralai/voxtral-small-24b-2507 | 32.768K | $0.1 | $0.3 | — | |||
| Google Gemini Flash Latest~google/gemini-flash-latest | 1.04858M | $0.75 | $3.75 | — | |||
| Google: Gemini 3.5 Flash (batch)google/gemini-3.5-flash:batch | 1.04858M | $0.75 | $4.5 | — | |||
| Google Gemini Pro Latest~google/gemini-pro-latest | 1.04858M | $2 | $12 | — | |||
| OpenAI: GPT Audio Miniopenai/gpt-audio-mini | 128K | $0.6 | $2.4 | — | |||
| NVIDIA: Nemotron 3 Nano Omni (free)nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free | 256K | Free | Free | — | |||
| Google: Gemini 2.5 Pro Preview 06-05google/gemini-2.5-pro-preview | 1.04858M | $1.25 | $10 | — | |||
| Google: Gemini 2.5 Pro (batch)google/gemini-2.5-pro:batch | 1.04858M | $0.625 | $5 | — | |||
| Google: Gemini 2.5 Flash Lite (batch)google/gemini-2.5-flash-lite:batch | 1.04858M | $0.05 | $0.2 | — | |||
| OpenAI: GPT Audioopenai/gpt-audio | 128K | $2.5 | $10 | — |