Google logo

Google: Gemini 3.1 Flash Lite

Low-latency Gemini model for high-volume multimodal and agent workloads

Source-linked gemini-flash-lite Closed weights Released 2026-05-07
API record Report
InputT
OutputT
Input price$0.25/M
Output price$1.5/M
Context1.04858M
Max output65.536K
Providers24
Inference availability

Providers

Provider-specific identifiers, limits, and listed prices per million tokens. Every row links back to the provider's own documentation.

Report a price
Providers offering Gemini 3.1 Flash Lite
ProviderProvider model IDContextMax outputInputOutputCache readCapabilitiesDocs
Kenari gemini-3-1-flash-lite 1.04858M 65.536K ReasoningToolsJSON Docs ↗
Kilo Gateway google/gemini-3.1-flash-lite 1.04858M 65.536K $0.125 $0.75 $0.013 ReasoningToolsJSON Docs ↗
Ofox google/gemini-3.1-flash-lite 1.04858M 65.536K $0.25 $1.5 $0.025 ReasoningToolsJSON Docs ↗
302.AI gemini-3.1-flash-lite 1.04858M 65.536K $0.25 $1.5 ReasoningToolsJSON Docs ↗
Requesty gemini-3.1-flash-lite 1.04858M 65.535K $0.25 $1.5 $0.025 ReasoningToolsJSON Docs ↗
Eden AI google/gemini-3.1-flash-lite 1.04858M 65.536K $0.25 $1.5 $0.025 ReasoningToolsJSON Docs ↗
Poe google/gemini-3.1-flash-lite 1.04858M 65.536K $0.25 $1.5 ReasoningTools Docs ↗
NanoGPT google/gemini-3.1-flash-lite 1.04858M 65.536K $0.25 $1.5 $0.025 ReasoningToolsJSON Docs ↗
NEAR AI Cloud google/gemini-3.1-flash-lite 1.04858M 65.536K $0.25 $1.5 $0.025 ReasoningToolsJSON Docs ↗
DevPass (LLM Gateway) gemini-3.1-flash-lite 1.04858M 65.536K $0.25 $1.5 $0.025 ReasoningToolsJSON Docs ↗
ZenMux google/gemini-3.1-flash-lite 1.04858M 65.536K $0.25 $1.5 $0.025 ReasoningToolsJSON Docs ↗
Abacus gemini-3.1-flash-lite 1.04858M 65.536K $0.25 $1.5 $0.025 ReasoningToolsJSON Docs ↗
AIHubMix gemini-3.1-flash-lite 1.04858M 65.536K $0.25 $1.5 $0.025 ReasoningToolsJSON Docs ↗
Vertex gemini-3.1-flash-lite 1.04858M 65.536K $0.25 $1.5 $0.025 ReasoningToolsJSON Docs ↗
Merge Gateway google/gemini-3.1-flash-lite 1.04858M 65.536K $0.25 $1.5 $0.025 ReasoningToolsJSON Docs ↗
Vercel AI Gateway google/gemini-3.1-flash-lite 1M 65K $0.25 $1.5 $0.03 ReasoningToolsJSON Docs ↗
OpenRouter google/gemini-3.1-flash-lite 1.04858M 65.536K $0.25 $1.5 $0.025 ReasoningToolsJSON Docs ↗
Google gemini-3.1-flash-lite 1.04858M 65.536K $0.25 $1.5 $0.025 ReasoningToolsJSON Docs ↗
Impossibl google/gemini-3.1-flash-lite 1.04858M 65.536K $0.25 $1.5 $0.025 ReasoningToolsJSON Docs ↗
OrcaRouter google/gemini-3.1-flash-lite 1.04858M 65.536K $0.25 $1.5 $0.025 ReasoningToolsJSON Docs ↗
Eden AI vertex/gemini-3.1-flash-lite 1.04858M 65.536K $0.25 $1.5 $0.025 ReasoningToolsJSON Docs ↗
Cloudflare AI Gateway google-ai-studio/gemini-3.1-flash-lite 1.04858M 65.536K $0.25 $1.5 $0.025 ReasoningToolsJSON Docs ↗
Pioneer gemini-3.1-flash-lite 1M 65K $0.25 $1.5 $0.03 ReasoningToolsJSON Docs ↗
Cortecs gemini-3.1-flash-lite 1.04858M 65.535K $0.272 $1.631 $0.025 ReasoningToolsJSON Docs ↗

Capability badges appear only where the provider catalog explicitly lists support. A blank cell means the source is silent, not that the feature is absent.

Listed rates

Price across providers

Input price per million tokens as published by each provider. Bars are drawn from listed rates only — no traffic weighting, since the catalog observes no requests.

Lowest input $0.125/M

Across 23 priced providers

Median input $0.25/M

Midpoint of listed rates

Highest input $0.272/M

2.2× the lowest listed rate

Output range $0.75 – $1.631

Per million output tokens

Kilo Gateway $0.125/MLowest
Ofox $0.25/M
302.AI $0.25/M
Requesty $0.25/M
Eden AI $0.25/M
Poe $0.25/M
NanoGPT $0.25/M
NEAR AI Cloud $0.25/M
DevPass (LLM Gateway) $0.25/M
ZenMux $0.25/M
Abacus $0.25/M
AIHubMix $0.25/M
Vertex $0.25/M
Merge Gateway $0.25/M
Vercel AI Gateway $0.25/M
OpenRouter $0.25/M
Google $0.25/M
Impossibl $0.25/M
OrcaRouter $0.25/M
Eden AI $0.25/M
Cloudflare AI Gateway $0.25/M
Pioneer $0.25/M
Cortecs $0.272/M
Cost calculator

Estimate a workload

$0.00
Excludes taxes, non-token charges, and tiered discounts.

Context limits also differ by provider, from 1M to 1.04858M tokens. Compare the provider table above before choosing on price alone.

Specification

Capabilities

Recorded from the source catalog and provider listings.

Reasoning Yes
Tool calling Yes
Structured output Yes
Attachments Yes
Vision input Yes
× Open weights No
Creator
Google
Model family
gemini-flash-lite
Knowledge cutoff
2025-01
License
Not documented
Release date
2026-05-07
Model ID
google/gemini-3.1-flash-lite
Published evaluations

Benchmarks

Every result stays attached to its source, version, metric and harness. Scores from different versions are never merged, and the profile below plots each benchmark against its own population rather than on a shared scale.

Benchmark registry
No published benchmark results

ModelBench does not infer quality from price, context size, or model name. When a source publishes a comparable result, it appears here with its version and link.

Read the methodology
Catalog activity

Change log

Field-level changes detected between successful source imports.

Full change log
No changes recorded

This record has not changed within the retained import history.

Provenance

Sources & verification

Every figure on this page traces back to one of these records.

Methodology
Public API

Use this record

Fetch the complete source-linked model record. No key, no account, no rate-limited tier.

API documentation
Endpoint
GET https://model.kyssta.lol/api/v1/models/google/gemini-3.1-flash-lite
curl
curl "https://model.kyssta.lol/api/v1/models/google/gemini-3.1-flash-lite"
Common questions

Frequently asked questions

Answered directly from the stored record — nothing here is generated beyond the catalog's own fields.

What is Gemini 3.1 Flash Lite?

Low-latency Gemini model for high-volume multimodal and agent workloads. It is published by Google and catalogued here from Models.dev.

How much does Gemini 3.1 Flash Lite cost?

Listed input pricing starts at $0.125 per million tokens from Kilo Gateway, rising to $0.272 across 23 listed providers.

What is the context length of Gemini 3.1 Flash Lite?

Gemini 3.1 Flash Lite accepts up to 1.04858M tokens of context and returns up to 65.536K output tokens.

Does Gemini 3.1 Flash Lite support tool calling and structured output?

Provider catalogs list support for tool calling, structured output, reasoning, and image input.

Which providers serve Gemini 3.1 Flash Lite?

23 providers list this model: Kenari, Kilo Gateway, Ofox, 302.AI, Requesty, Eden AI and 17 more.

Are the weights for Gemini 3.1 Flash Lite open?

No. This model is served through hosted APIs only.

When was Gemini 3.1 Flash Lite released?

The catalog records a release date of 2026-05-07, last verified Sep 11, 2026.

More models from Google

View all →