Google logo

Google: Gemini 2.5 Flash-Lite

Lean Gemini 2.5 lane for cheap multimodal traffic and quick agents

Source-linked gemini-flash-lite Closed weights Released 2025-06-17
API record Report
InputT
OutputT
Input price$0.1/M
Output price$0.4/M
Context1.04858M
Max output65.536K
Providers18
Inference availability

Providers

Provider-specific identifiers, limits, and listed prices per million tokens. Every row links back to the provider's own documentation.

Report a price
Providers offering Gemini 2.5 Flash-Lite
ProviderProvider model IDContextMax outputInputOutputCache readCapabilitiesDocs
Kenari gemini-2-5-flash-lite 1.04858M 65.536K ReasoningToolsJSON Docs ↗
AnyAPI google/gemini-2.5-flash-lite 1.04858M 65.536K ReasoningToolsJSON Docs ↗
Poe google/gemini-2.5-flash-lite 1.024M 64K $0.07 $0.28 ReasoningTools Docs ↗
CrossModel gemini/gemini-2.5-flash-lite 1.04858M 65.536K $0.1 $0.4 $0.01 ReasoningToolsJSON Docs ↗
Google gemini-2.5-flash-lite 1.04858M 65.536K $0.1 $0.4 $0.01 ReasoningToolsJSON Docs ↗
NEAR AI Cloud google/gemini-2.5-flash-lite 1.04858M 65.536K $0.1 $0.4 $0.01 ReasoningToolsJSON Docs ↗
DevPass (LLM Gateway) gemini-2.5-flash-lite 1.04858M 65.536K $0.1 $0.4 $0.01 ReasoningToolsJSON Docs ↗
Ofox google/gemini-2.5-flash-lite 1.04858M 65.536K $0.1 $0.4 $0.01 ReasoningToolsJSON Docs ↗
ZenMux google/gemini-2.5-flash-lite 1.048M 64K $0.1 $0.4 $0.03 Tools Docs ↗
OrcaRouter google/gemini-2.5-flash-lite 1.04858M 65.536K $0.1 $0.4 $0.01 ReasoningToolsJSON Docs ↗
Vertex gemini-2.5-flash-lite 1.04858M 65.536K $0.1 $0.4 $0.01 ReasoningTools Docs ↗
Kilo Gateway google/gemini-2.5-flash-lite 1.04858M 65.535K $0.1 $0.4 $0.01 ReasoningToolsJSON Docs ↗
Merge Gateway google/gemini-2.5-flash-lite 1M 65.536K $0.1 $0.4 $0.01 ReasoningToolsJSON Docs ↗
Vercel AI Gateway google/gemini-2.5-flash-lite 1.04858M 65.536K $0.1 $0.4 $0.01 ReasoningToolsJSON Docs ↗
Cloudflare AI Gateway google-ai-studio/gemini-2.5-flash-lite 1.04858M 65.536K $0.1 $0.4 $0.01 ReasoningToolsJSON Docs ↗
OpenRouter google/gemini-2.5-flash-lite 1.04858M 65.535K $0.1 $0.4 $0.01 ReasoningToolsJSON Docs ↗
LLMTR google/gemini-2.5-flash-lite 1.04858M 65.536K $0.1 $0.1 ReasoningToolsJSON Docs ↗
Impossibl google/gemini-2.5-flash-lite 1.04858M 65.536K $0.1 $0.4 $0.01 ReasoningToolsJSON Docs ↗

Capability badges appear only where the provider catalog explicitly lists support. A blank cell means the source is silent, not that the feature is absent.

Listed rates

Price across providers

Input price per million tokens as published by each provider. Bars are drawn from listed rates only — no traffic weighting, since the catalog observes no requests.

Lowest input $0.07/M

Across 16 priced providers

Median input $0.1/M

Midpoint of listed rates

Highest input $0.1/M

1.4× the lowest listed rate

Output range $0.1 – $0.4

Per million output tokens

Poe $0.07/MLowest
CrossModel $0.1/M
Google $0.1/M
NEAR AI Cloud $0.1/M
DevPass (LLM Gateway) $0.1/M
Ofox $0.1/M
ZenMux $0.1/M
OrcaRouter $0.1/M
Vertex $0.1/M
Kilo Gateway $0.1/M
Merge Gateway $0.1/M
Vercel AI Gateway $0.1/M
Cloudflare AI Gateway $0.1/M
OpenRouter $0.1/M
LLMTR $0.1/M
Impossibl $0.1/M
Cost calculator

Estimate a workload

$0.00
Excludes taxes, non-token charges, and tiered discounts.

Context limits also differ by provider, from 1M to 1.04858M tokens. Compare the provider table above before choosing on price alone.

Specification

Capabilities

Recorded from the source catalog and provider listings.

Reasoning Yes
Tool calling Yes
Structured output Yes
Attachments Yes
Vision input Yes
× Open weights No
Creator
Google
Model family
gemini-flash-lite
Knowledge cutoff
2025-01
License
Not documented
Release date
2025-06-17
Model ID
google/gemini-2.5-flash-lite
Published evaluations

Benchmarks

Every result stays attached to its source, version, metric and harness. Scores from different versions are never merged, and the profile below plots each benchmark against its own population rather than on a shared scale.

Benchmark registry
Benchmark profile
SciCodeSciCode AA Coding IndexArtificial Analysis Coding Index Terminal-Bench…Terminal-Bench Hard This model — SciCode: 19.3 percent correct (1.9th percentile) This model — Artificial Analysis Coding Index: 9.5 index (1.9th percentile) This model — Terminal-Bench Hard: 4.5 success rate (11.4th percentile)

Each axis is this model's percentile among the 3 benchmarks it has published results for, measured against every other model with a score on that same benchmark. Percentiles are used because benchmarks do not share a scale — a 60 on one is not a 60 on another. Hover any point for the raw score.

SciCode percent correct · 2026-03-11 Source ↗
SciCode #27 of 27
Artificial Analysis Coding Index index · 2026-03-11 Source ↗
Artificial Analysis Coding Index #26 of 26
Terminal-Bench Hard success rate · 2026-03-11 Source ↗
Terminal-Bench Hard #20 of 22

Each strip shows every published score for that benchmark, with this model marked. The lighter shape behind the ticks is the density of results, and the dashed line is the median.

Cost against capability

Price and performance

Listed input price plotted against SciCode, the benchmark with the widest published coverage that this model appears in.

0 23 45 $0.1 $1 $10 Qwen3-Coder 30B-A3B Instruct — $0.45/M, score 27.8 Qwen3 Max — $1.2/M, score 38.3 DeepSeek-R1 — $0.7/M, score 35.7 Gemini 2.5 Flash — $0.3/M, score 39.4 Gemini 2.5 Flash-Lite — $0.1/M, score 19.3 Gemini 2.5 Pro — $1.25/M, score 42.8 Llama-3.3-70B-Instruct — $0.1/M, score 26.0 Devstral 2 — $0.4/M, score 33.1 Mistral Large 2.1 — $2/M, score 29.2 Mistral Large 3 — $0.5/M, score 36.2 Mistral Medium 3 — $0.4/M, score 33.1 Mistral Small 4 — $0.15/M, score 38.0 GPT-4 Turbo — $10/M, score 31.9 GPT-4o (2024-05-13) — $5/M, score 30.9 GPT-4o (2024-08-06) — $2.5/M, score 33.1 GPT-4o (2024-11-20) — $2.5/M, score 33.3 GPT-4o mini — $0.15/M, score 22.9 GPT-5-Codex — $1.1/M, score 40.9 Sonar — $1/M, score 22.9 Sonar Pro — $3/M, score 22.6 Step 3.5 Flash — $0.1/M, score 40.4 Step 3.5 Flash 2603 — $0.1/M, score 38.5 Step 3.7 Flash — $0.185/M, score 40.0 GLM-4.5 — $0.6/M, score 34.8 GLM-4.5-Air — $0.2/M, score 30.6 GLM-4.5V — $0.6/M, score 22.1 GLM-4.6 — $0.6/M, score 38.4 Gemini 2.5 Flash-Lite Input price per million tokens (log scale) Percent correct

The stepped line is the efficient frontier: at each price, the best score available for that money or less. A model sitting on it is not being beaten by anything cheaper. Price is log-scaled because listed rates span four orders of magnitude. Only models with both a listed price and a score on this benchmark can appear.

Catalog activity

Change log

Field-level changes detected between successful source imports.

Full change log
Max Output Tokens65535 → 65536
Price Completion0.39999999999999997 → 0.4
Price Prompt0.09999999999999999 → 0.1
Max Output Tokens65536 → 65535
Price Completion0.4 → 0.39999999999999997
Price Prompt0.1 → 0.09999999999999999
Max Output Tokens65535 → 65536
Price Completion0.39999999999999997 → 0.4
Price Prompt0.09999999999999999 → 0.1
Max Output Tokens65536 → 65535
Provenance

Sources & verification

Every figure on this page traces back to one of these records.

Methodology
Public API

Use this record

Fetch the complete source-linked model record. No key, no account, no rate-limited tier.

API documentation
Endpoint
GET https://model.kyssta.lol/api/v1/models/google/gemini-2.5-flash-lite
curl
curl "https://model.kyssta.lol/api/v1/models/google/gemini-2.5-flash-lite"
Common questions

Frequently asked questions

Answered directly from the stored record — nothing here is generated beyond the catalog's own fields.

What is Gemini 2.5 Flash-Lite?

Lean Gemini 2.5 lane for cheap multimodal traffic and quick agents. It is published by Google and catalogued here from Models.dev.

How much does Gemini 2.5 Flash-Lite cost?

Listed input pricing starts at $0.07 per million tokens from Poe, rising to $0.1 across 16 listed providers.

What is the context length of Gemini 2.5 Flash-Lite?

Gemini 2.5 Flash-Lite accepts up to 1.04858M tokens of context and returns up to 65.536K output tokens.

Does Gemini 2.5 Flash-Lite support tool calling and structured output?

Provider catalogs list support for tool calling, structured output, reasoning, and image input.

Which providers serve Gemini 2.5 Flash-Lite?

18 providers list this model: Kenari, AnyAPI, Poe, CrossModel, Google, NEAR AI Cloud and 12 more.

Are the weights for Gemini 2.5 Flash-Lite open?

No. This model is served through hosted APIs only.

When was Gemini 2.5 Flash-Lite released?

The catalog records a release date of 2025-06-17, last verified Sep 11, 2026.

More models from Google

View all →