Reasoning Tools JSON

Google's most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows

google/gemini-3.8-flash 2026-09-02 1.04858M context $0.75/M input $3.75/M output
17 providers
Reasoning Tools JSON

High-efficiency Gemini model for agentic workflows, coding, and multimodal reasoning

google/gemini-3.7-flash 2026-08-13 1.04858M context $0.75/M input $3.75/M output
23 providers
Reasoning Tools JSON

High-efficiency Gemini model for agentic workflows, coding, and multimodal reasoning

google/gemini-flash-latest 2026-08-13 1.04858M context $0.75/M input $3.75/M output
7 providers
Reasoning Tools JSON

Fast Gemini model balancing multimodal reasoning, tool use, and cost

google/gemini-flash-lite-latest 2026-07-21 1.04858M context $0.3/M input $2.5/M output
6 providers
Reasoning Tools JSON

Fast Gemini model balancing multimodal reasoning, tool use, and cost

google/gemini-3.5-flash-lite 2026-07-21 1.04858M context $0.3/M input $2.5/M output
23 providers
Reasoning Tools JSON

Fast Gemini model balancing multimodal reasoning, tool use, and cost

google/gemini-3.6-flash 2026-07-21 1.04858M context $0.75/M input $3.75/M output
24 providers
Reasoning Tools

Fastest, most cost-efficient Gemini image model for high-volume 1K generation and editing

google/gemini-3.1-flash-lite-image 2026-06-30 65.536K context $0.25/M input $30/M output
5 providers
Reasoning

Video generation and editing model for fast, conversational text- and image-to-video workflows

google/gemini-omni-flash-preview 2026-06-30 1.04858M context $1.5/M input $17.5/M output
2 providers

Low-latency audio-to-audio model for real-time speech translation across 70+ languages

google/gemini-3.5-live-translate-preview 2026-06-09 131.072K context $3.5/M input $21/M output
1 provider
Reasoning

Image model for prompt-driven generation, editing, and visual design workflows

google/gemini-3.1-flash-image 2026-05-28 131.072K context $0.5/M input $60/M output
9 providers
Reasoning

Nano Banana Pro for higher-fidelity image generation and design-heavy edits

google/gemini-3-pro-image 2026-05-28 65.536K context $2/M input $120/M output
9 providers
Reasoning Tools JSON

Fast Gemini model balancing multimodal reasoning, tool use, and cost

google/gemini-3.5-flash 2026-05-19 1.04858M context $1.5/M input $9/M output
31 providers
Reasoning Tools JSON

Low-latency Gemini model for high-volume multimodal and agent workloads

google/gemini-3.1-flash-lite 2026-05-07 1.04858M context $0.25/M input $1.5/M output
23 providers

Multimodal embedding model mapping text, images, video, audio, and PDFs into a unified embedding space

google/gemini-embedding-2 2026-04-22 8.192K context $0.2/M input Output not listed
3 providers

Agentic model for autonomous multi-step research, synthesis, and cited reports

google/deep-research-preview-04-2026 2026-04-21 1.04858M context Input not listed Output not listed
1 provider

Maximum-comprehensiveness agentic researcher for multi-step investigation, synthesis, and cited reports

google/deep-research-max-preview-04-2026 2026-04-21 1.04858M context Input not listed Output not listed
1 provider

Low-latency speech generation with steerable prompts and expressive audio tags

google/gemini-3.1-flash-tts-preview 2026-04-15 8.192K context $1/M input $20/M output
2 providers
Reasoning Tools JSON

Vision-language model for embodied reasoning: spatial understanding, task planning, and physical-world agentic robotics

google/gemini-robotics-er-1.6-preview 2026-04-14 131.072K context $1/M input $5/M output
2 providers
Reasoning Tools

High-quality, low-latency Live API model for real-time dialogue and voice-first AI applications

google/gemini-3.1-flash-live-preview 2026-03-26 131.072K context $0.75/M input $4.5/M output
1 provider

Music generation model for full-length songs from text or images with vocals and structure

google/lyria-3-pro-preview 2026-03-25 131.072K context Input not listed Output not listed
3 providers

Music generation model for short 30-second clips, loops, and previews from text or image prompts

google/lyria-3-clip-preview 2026-03-25 131.072K context Input not listed Output not listed
4 providers
Reasoning Tools JSON

Low-latency Gemini model for high-volume multimodal and agent workloads

google/gemini-3.1-flash-lite-preview 2026-03-03 1.04858M context $0.25/M input $1.5/M output
13 providers
Reasoning

Image model for prompt-driven generation, editing, and visual design workflows

google/gemini-3.1-flash-image-preview 2026-02-26 65.536K context $0.5/M input $3/M output
9 providers
Reasoning Tools JSON

Advanced Gemini model for complex reasoning, coding, and multimodal analysis

google/gemini-3.1-pro-preview-customtools 2026-02-19 1.04858M context $2/M input $12/M output
11 providers
Reasoning Tools JSON

Reasoning-first Gemini preview for agentic coding and complex problem solving

google/gemini-3.1-pro-preview 2026-02-19 1.04858M context $2/M input $12/M output
31 providers
Reasoning Tools JSON

New Gemini flash lane bringing frontier-style multimodal reasoning to cheaper runs

google/gemini-3-flash-preview 2025-12-17 1.04858M context $0.5/M input $3/M output
25 providers
Reasoning

Nano Banana Pro for higher-fidelity image generation and design-heavy edits

google/gemini-3-pro-image-preview 2025-11-20 65.536K context $1/M input $6/M output
8 providers
Reasoning Tools JSON

Preview Gemini flagship for complex reasoning, coding, and rich multimodal prompts

google/gemini-3-pro-preview 2025-11-18 1.04858M context $0.57/M input $3.43/M output
10 providers
Reasoning Tools JSON

Specialized Gemini 2.5 model for browser-control agents that automate UI tasks

google/gemini-2.5-computer-use-preview-10-2025 2025-10-07 128K context $1.25/M input $10/M output
2 providers

Speech generation model for controllable voice, narration, and audio delivery

google/gemini-2.5-flash-tts 2025-09-30 32.768K context $0.5/M input $10/M output
1 provider

Speech generation model for controllable voice, narration, and audio delivery

google/gemini-2.5-pro-tts 2025-09-30 32.768K context $1/M input $20/M output
1 provider

Nano Banana image model for fast generation, edits, and character-consistent assets

google/gemini-2.5-flash-image 2025-08-26 32.768K context $0.3/M input $30/M output
8 providers
Reasoning Tools JSON

Google's proven reasoning model for coding, math, and multimodal analysis

google/gemini-2.5-pro 2025-06-17 1.04858M context $1.25/M input $10/M output
29 providers
Reasoning Tools JSON

Fast Gemini workhorse for multimodal apps where latency and price matter

google/gemini-2.5-flash 2025-06-17 1.04858M context $0.3/M input $2.5/M output
30 providers
Reasoning Tools JSON

Lean Gemini 2.5 lane for cheap multimodal traffic and quick agents

google/gemini-2.5-flash-lite 2025-06-17 1.04858M context $0.1/M input $0.4/M output
18 providers
Tools

Low-latency Gemini model for high-volume multimodal and agent workloads

google/gemini-2.0-flash-lite 2024-12-11 1.04858M context $0.052/M input $0.21/M output
2 providers
Tools

Earlier Gemini Flash workhorse for responsive multimodal apps and tool use

google/gemini-2.0-flash 2024-12-11 1.04858M context $0.1/M input $0.42/M output
2 providers

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

google/gemini-3.7-flash:batch 1.04858M context $0.375/M input $1.875/M output

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

google/gemini-2.5-pro-preview 1.04858M context $1.25/M input $10/M output

Gemma 2 27B by Google is an open model built from the same research and technology used to create the [Gemini models](/models?q=gemini). Gemma models are well-suited for a variety of...

google/gemma-2-27b-it 8.192K context $0.65/M input $0.65/M output

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...

google/gemma-4-26b-a4b-it:free 262.144K context Free input Free output

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

google/gemma-4-31b-it:free 262.144K context Free input Free output

Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool...

google/gemini-3-flash-preview:batch 1.04858M context $0.25/M input $1.5/M output

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

google/gemini-3.5-flash:batch 1.04858M context $0.75/M input $4.5/M output

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

google/gemini-2.5-flash-lite:batch 1.04858M context $0.05/M input $0.2/M output

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

google/gemma-4-31b-it:batch 262.144K context $0.39/M input $0.97/M output

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...

google/gemini-2.5-flash:batch 1.04858M context $0.15/M input $1.25/M output

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

google/gemini-2.5-pro-preview-05-06 1.04858M context $1.25/M input $10/M output

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

google/gemini-3.5-flash-lite:batch 1.04858M context $0.15/M input $1.25/M output

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

google/gemini-3.8-flash:batch 1.04858M context $0.375/M input $1.875/M output