Inference availability
Providers Provider-specific identifiers, limits, and listed prices per million tokens. Every row links back to the provider's own documentation.
Report a price
Providers offering Gemini 3.5 Flash Lite
Provider Provider model ID Context Max output Input Output Cache read Capabilities Docs
S SAP AI Core
gemini-3.5-flash-lite
1.04858M
65.536K
—
—
—
Reasoning Tools JSON
Docs ↗
K Kilo Gateway
google/gemini-3.5-flash-lite
1.04858M
65.536K
$0.15
$1.25
$0.015
Reasoning Tools JSON
Docs ↗
E Eden AI
google/gemini-3.5-flash-lite
1.04858M
65.536K
$0.3
$2.5
$0.03
Reasoning Tools JSON
Docs ↗
I Impossibl
google/gemini-3.5-flash-lite
1.04858M
65.536K
$0.3
$2.5
$0.03
Reasoning Tools JSON
Docs ↗
E Eden AI
vertex/gemini-3.5-flash-lite
1.04858M
65.536K
$0.3
$2.5
$0.03
Reasoning Tools JSON
Docs ↗
G Google
gemini-3.5-flash-lite
1.04858M
65.536K
$0.3
$2.5
$0.03
Reasoning Tools JSON
Docs ↗
N NanoGPT
google/gemini-3.5-flash-lite
1.04858M
65.536K
$0.3
$2.5
$0.03
Reasoning Tools JSON
Docs ↗
V Vertex
gemini-3.5-flash-lite
1.04858M
65.536K
$0.3
$2.5
$0.03
Reasoning Tools JSON
Docs ↗
3 302.AI
gemini-3.5-flash-lite
1.04858M
65.536K
$0.3
$2.5
—
Reasoning Tools JSON
Docs ↗
A Abacus
gemini-3.5-flash-lite
1.04858M
65.536K
$0.3
$2.5
$0.03
Reasoning Tools JSON
Docs ↗
M Merge Gateway
google/gemini-3.5-flash-lite
1.04858M
65.536K
$0.3
$2.5
$0.03
Reasoning Tools JSON
Docs ↗
V Vercel AI Gateway
google/gemini-3.5-flash-lite
1M
65K
$0.3
$2.5
$0.03
Reasoning Tools JSON
Docs ↗
O OpenRouter
google/gemini-3.5-flash-lite
1.04858M
65.536K
$0.3
$2.5
$0.03
Reasoning Tools JSON
Docs ↗
O Opper
gemini/gemini-3.5-flash-lite
1.04858M
65.536K
$0.3
$2.5
$0.03
Reasoning Tools JSON
Docs ↗
O Ofox
google/gemini-3.5-flash-lite
1.04858M
65.536K
$0.3
$2.5
$0.03
Reasoning Tools JSON
Docs ↗
D DevPass (LLM Gateway)
gemini-3.5-flash-lite
1.04858M
65.536K
$0.3
$2.5
$0.03
Reasoning Tools JSON
Docs ↗
O OpenCode Zen
gemini-3.5-flash-lite
1.04858M
65.536K
$0.3
$2.5
$0.03
Reasoning Tools JSON
Docs ↗
P Pioneer
gemini-3.5-flash-lite
1M
65K
$0.3
$2.5
$0.03
Reasoning Tools JSON
Docs ↗
C CrossModel
gemini/gemini-3.5-flash-lite
1.04858M
65.536K
$0.3
$2.5
$0.03
Reasoning Tools JSON
Docs ↗
N Neon
gemini-3-5-flash-lite
1.04858M
65.536K
$0.3
$2.5
$0.03
Reasoning Tools JSON
Docs ↗
R Requesty
gemini-3.5-flash-lite
1.04858M
65.535K
$0.3
$2.5
$0.03
Reasoning Tools JSON
Docs ↗
O OrcaRouter
google/gemini-3.5-flash-lite
1.04858M
65.536K
$0.3
$2.5
$0.03
Reasoning Tools JSON
Docs ↗
C Cloudflare AI Gateway
google-ai-studio/gemini-3.5-flash-lite
1.04858M
65.536K
$0.3
$2.5
$0.03
Reasoning Tools JSON
Docs ↗
C Cortecs
gemini-3.5-flash-lite
1.04858M
65.535K
$0.33
$2.749
$0.033
Reasoning Tools JSON
Docs ↗
Capability badges appear only where the provider catalog explicitly lists support. A blank cell means the source is silent, not that the feature is absent.
Listed rates
Price across providers Input price per million tokens as published by each provider. Bars are drawn from listed rates only — no traffic weighting, since the catalog observes no requests.
Lowest input
$0.15/M
Across 23 priced providers
Median input
$0.3/M
Midpoint of listed rates
Highest input
$0.33/M
2.2× the lowest listed rate
Output range
$1.25 – $2.749
Per million output tokens
Kilo Gateway
$0.15/M Lowest
Eden AI
$0.3/M
Impossibl
$0.3/M
Eden AI
$0.3/M
Google
$0.3/M
NanoGPT
$0.3/M
Vertex
$0.3/M
302.AI
$0.3/M
Abacus
$0.3/M
Merge Gateway
$0.3/M
Vercel AI Gateway
$0.3/M
OpenRouter
$0.3/M
Opper
$0.3/M
Ofox
$0.3/M
DevPass (LLM Gateway)
$0.3/M
OpenCode Zen
$0.3/M
Pioneer
$0.3/M
CrossModel
$0.3/M
Neon
$0.3/M
Requesty
$0.3/M
OrcaRouter
$0.3/M
Cloudflare AI Gateway
$0.3/M
Cortecs
$0.33/M
Context limits also differ by provider, from 1M to 1.04858M tokens. Compare the provider table above before choosing on price alone.
Specification
Capabilities Recorded from the source catalog and provider listings.
✓
Reasoning
Yes
✓
Tool calling
Yes
✓
Structured output
Yes
✓
Attachments
Yes
✓
Vision input
Yes
×
Open weights
No
Model family gemini-flash-lite
Knowledge cutoff 2026-03
License Not documented
Release date 2026-07-21
Model ID google/gemini-3.5-flash-lite
Published evaluations
Benchmarks Every result stays attached to its source, version, metric and harness. Scores from different versions are never merged, and the profile below plots each benchmark against its own population rather than on a shared scale.
Benchmark registry
Benchmark profile
Coding IndexCoding Index
Agentic IndexAgentic Index
IntelligenceIntelligence Index
SWE-Bench ProSWE-Bench Pro
OSWorld-VerifiedOSWorld-Verified
CharXiv Reasoni…CharXiv Reasoning
GDPval-AAGDPval-AA v2
This model — Coding Index: 49.3 index (51.0th percentile)
This model — Agentic Index: 15.9 index (42.2th percentile)
This model — Intelligence Index: 22.7 index (41.1th percentile)
This model — SWE-Bench Pro: 54.2 resolve rate (47.4th percentile)
This model — OSWorld-Verified: 74 success rate (26.5th percentile)
This model — CharXiv Reasoning: 76.5 accuracy (7.1th percentile)
This model — GDPval-AA v2: 1140 Elo (12.5th percentile)
Each axis is this model's percentile among the 7 benchmarks it has published results for, measured against every other model with a score on that same benchmark. Percentiles are used because benchmarks do not share a scale — a 60 on one is not a 60 on another. Hover any point for the raw score.
Coding Index
#99 of 205
2.7
median 49.3
81.6
Agentic Index
#111 of 193
0.1
median 19.7
58.0
Intelligence Index
#113 of 192
3.8
median 25.6
55.0
SWE-Bench Pro
#20 of 38
5.2
median 54.3
80.3
OSWorld-Verified
#13 of 17
39.0
median 78.4
86.1
CharXiv Reasoning
#7 of 7
76.5
median 84.2
91.3
GDPval-AA
#4 of 4
1140.0
median 1544.5
1861.0
GDM-MRCR
#3 of 3
21.3
median 54.0
97.0
39.2
2 models scored — too few for a distribution
54.0
2 models scored — too few for a distribution
Each strip shows every published score for that benchmark, with this model marked. The lighter shape behind the ticks is the density of results, and the dashed line is the median.
Cost against capability
Price and performance Listed input price plotted against Coding Index, the benchmark with the widest published coverage that this model appears in.
0
43
86
$0.01
$0.1
$1
$10
SpaceXAI: Grok 4.3 (batch) — $1/M, score 42.2
Thinking Machines: Inkling (batch) — $1/M, score 52.1
Kimi K2.6 — $0.95/M, score 61.8
SpaceXAI: Grok 4.5 — $2/M, score 72.4
GPT-5.4 — $2.5/M, score 71.1
Anthropic: Claude Sonnet 4.5 — $3/M, score 52.1
Mistral: Mistral Large 3 2512 (batch) — $0.25/M, score 20.1
GPT-5 — $1.25/M, score 37.8
OpenAI: GPT-3.5 Turbo (batch) — $0.25/M, score 10.7
OpenAI: GPT-4 Turbo (batch) — $5/M, score 21.5
OpenAI: o3 Mini High (batch) — $0.55/M, score 16.3
Qwen: Qwen3.8 Max (0902) — $2/M, score 71.8
Z.ai: GLM 5.3 Flash (batch) — $0.075/M, score 71.5
Nemotron 3 Super 120B A12B — $0.2/M, score 37.7
Claude Opus 5 — $5/M, score 78.0
Claude Sonnet 5 — $2/M, score 71.5
Mistral: Ministral 3 8B 2512 — $0.15/M, score 9.7
OpenAI: o1 (batch) — $7.5/M, score 39.7
Gemini 3.7 Flash — $0.75/M, score 76.1
LongCat-2.0 — $0.3/M, score 45.3
Qwen: Qwen3.7 Plus — $0.32/M, score 55.9
Mistral: Mistral Large 3 2512 — $0.5/M, score 20.1
Google: Gemini 3.5 Flash (batch) — $0.75/M, score 70.1
OpenAI: GPT-5.4 Nano (batch) — $0.1/M, score 56.1
MiMo-V2.5-Pro — $0.435/M, score 60.2
GPT-4 — $30/M, score 13.1
OpenAI: GPT-4o-mini (batch) — $0.075/M, score 11.4
Qwen: Qwen3.5-9B (batch) — $0.17/M, score 28.7
Anthropic: Claude Opus 4.8 (batch) — $2.5/M, score 74.3
Gemini 3.6 Flash — $0.75/M, score 69.2
Kimi K3 — $3/M, score 76.2
Z.ai: GLM 5.1 — $0.966/M, score 55.8
Nemotron 3 Nano 30B A3B — $0.05/M, score 14.4
Gemma 4 26B A4B IT — $0.042/M, score 39.3
Claude Fable 5 — $10/M, score 76.5
Gemma 4 31B IT — $0.09/M, score 43.4
Anthropic: Claude Opus 4.7 — $5/M, score 73.6
DeepSeek V3.2 — $0.18/M, score 44.2
Qwen: Qwen3.6 Plus — $0.325/M, score 54.5
MiniMax: MiniMax M3 — $0.3/M, score 58.6
MiniMax: MiniMax M2.7 — $0.3/M, score 52.6
DeepSeek: DeepSeek V4 Pro 0813 (batch) — $0.66/M, score 68.8
Thinking Machines: Inkling Small (batch) — $0.5/M, score 52.9
Gemma 3 12B IT — $0.05/M, score 5.8
Anthropic: Claude Haiku 4.5 (batch) — $0.5/M, score 43.9
Anthropic: Claude Fable 5.1 — $10/M, score 81.6
Google: Gemma 4 31B (batch) — $0.39/M, score 43.4
Qwen: Qwen3.8 2.4T A95B (batch) — $2/M, score 71.9
Mistral: Devstral 2 2512 — $0.4/M, score 31.3
MoonshotAI: Kimi K2.7 Code (batch) — $0.95/M, score 60.8
Inkling — $1.87/M, score 52.1
OpenAI: gpt-oss-20b (batch) — $0.05/M, score 20.7
MiMo-V2.5 — $0.14/M, score 56.8
Google: Gemini 3.5 Flash Lite (batch) — $0.15/M, score 49.3
Kimi K2 Thinking — $0.4/M, score 21.0
Mistral: Mistral Medium 3.5 (batch) — $0.75/M, score 46.9
Qwen: Qwen3.8 Max (0803) — $2/M, score 68.9
Mistral: Mistral Medium 3.1 (batch) — $0.2/M, score 20.5
Muse Spark 1.1 — $1.25/M, score 71.3
GPT-5.1 — $1.25/M, score 49.4
Qwen: Qwen3.8 27B — $0.42/M, score 68.1
Google: Gemini 3.8 Flash (batch) — $0.375/M, score 76.3
Muse Spark 1.2 — $1.25/M, score 72.2
Mistral: Mistral Small 4 (batch) — $0.075/M, score 26.6
inclusionAI: Ling 3.0 Flash VL — $0.06/M, score 57.0
Z.ai: GLM 5.3 (batch) — $0.7/M, score 74.8
DeepSeek: DeepSeek V4 Flash 0731 (batch) — $0.11/M, score 69.1
MoonshotAI: Kimi K3 (batch) — $3/M, score 76.2
GPT-5.6 Sol — $4/M, score 77.4
Anthropic: Claude Fable 5 (batch) — $5/M, score 76.5
IBM: Granite 4.2 8B — $0.06/M, score 22.4
GPT-5.4 nano — $0.2/M, score 56.1
GPT-5 Mini — $0.25/M, score 15.6
Google: Gemini 2.5 Pro (batch) — $0.625/M, score 33.3
Anthropic: Claude Sonnet 5 (batch) — $1/M, score 71.5
Google: Gemini 3.7 Flash (batch) — $0.375/M, score 76.1
DeepSeek-R1 — $0.7/M, score 24.6
MiniMax: MiniMax M3 (batch) — $0.3/M, score 58.6
OpenAI: GPT-6 Astra (batch) — $5/M, score 76.9
OpenAI: GPT-5.6 Terra (batch) — $1/M, score 76.7
Gemini 3.5 Flash — $1.5/M, score 70.1
Z.ai: GLM 5.3 Flash — $0.15/M, score 71.5
Gemini 2.5 Pro — $1.25/M, score 33.3
Anthropic: Claude Sonnet 4.6 — $3/M, score 63.0
DeepSeek: DeepSeek V3.1 Terminus — $0.27/M, score 43.5
SpaceXAI: Grok 4.6 — $2/M, score 76.8
Google: Gemini 3.6 Flash (batch) — $0.375/M, score 69.2
DeepSeek V4 Pro 0813 — $0.442/M, score 68.8
Z.ai: GLM 5.2 (batch) — $0.7/M, score 68.8
OpenAI: GPT-5.4 (batch) — $1.25/M, score 71.1
OpenAI: gpt-oss-120b (batch) — $0.15/M, score 30.4
Z.ai: GLM 5.3 — $1.4/M, score 74.8
Anthropic: Claude Opus 4.8 — $5/M, score 74.3
SpaceXAI: Grok 4.3 — $1.25/M, score 42.2
Mistral: Ministral 3 8B 2512 (batch) — $0.075/M, score 9.7
OpenAI: GPT-5.1 (batch) — $0.625/M, score 49.4
GPT-5.6 Terra — $2/M, score 76.7
Anthropic: Claude Sonnet 4.5 (batch) — $1.5/M, score 52.1
Qwen: Qwen3 Next 80B A3B Thinking — $0.15/M, score 17.4
NVIDIA: Nemotron 3 Ultra (batch) — $0.6/M, score 49.3
Nemotron 3 Ultra 550B A55B — $0.5/M, score 49.3
Gemini 3.8 Flash — $0.75/M, score 76.3
Gemma 3 4B IT — $0.04/M, score 2.7
Anthropic: Claude Opus 4.7 (batch) — $2.5/M, score 73.6
Solar Pro 4 — $0.3/M, score 52.7
Gemini 3.1 Pro Preview — $2/M, score 68.8
Google: Gemini 3.1 Pro Preview (batch) — $1/M, score 68.8
Gemma 3 27B IT — $0.08/M, score 10.1
Mistral: Mistral Medium 3.1 — $0.4/M, score 20.5
Inkling Small — $0.45/M, score 52.9
OpenAI: GPT-5.5 (batch) — $2.5/M, score 74.9
GPT-4o (2024-05-13) — $5/M, score 24.2
Anthropic: Claude Sonnet 4.6 (batch) — $1.5/M, score 63.0
DeepSeek V4 Flash 0731 — $0.05/M, score 69.1
GPT OSS 120B — $0.03/M, score 30.4
DeepSeek V4 Flash — $0.15/M, score 52.0
inclusionAI: Ling 3.0 Flash — $0.021/M, score 50.6
DeepSeek V4 Pro — $0.435/M, score 59.4
OpenAI: GPT-5.6 Luna (batch) — $0.1/M, score 71.4
OpenAI: GPT-5.4 Mini (batch) — $0.375/M, score 56.1
Qwen: Qwen3.8 2.4T A95B — $2/M, score 71.9
OpenAI: GPT-5.6 Sol (batch) — $1/M, score 77.4
GPT-6 Astra — $10/M, score 76.9
GPT-4.1 mini — $0.4/M, score 20.2
GPT-4 Turbo — $10/M, score 21.5
o1 — $15/M, score 39.7
GPT-3.5-turbo — $0.5/M, score 10.7
GPT-5.6 Luna — $0.2/M, score 71.4
GPT-5.5 — $5/M, score 74.9
Kimi K2.7 Code — $0.95/M, score 60.8
Kimi K2.5 — $0.3/M, score 46.8
Gemini 3.5 Flash Lite — $0.3/M, score 49.3
Hy3 preview — $0.066/M, score 58.8
Anthropic: Claude Fable 5.1 (batch) — $5/M, score 81.6
Claude Opus 5 (batch) — $2.5/M, score 78.0
Gemini 3.1 Flash Lite Preview — $0.25/M, score 34.7
Inception: Mercury 2 — $0.25/M, score 31.1
Z.ai: GLM 4.6 — $0.43/M, score 45.8
Nemotron 3.5 Lightning 30B A3B — $0.05/M, score 26.8
OpenAI: GPT-5 (batch) — $0.625/M, score 37.8
OpenAI: GPT-5 Mini (batch) — $0.125/M, score 15.6
Anthropic: Claude Sonnet 4 — $3/M, score 37.6
OpenAI: GPT-4.1 Mini (batch) — $0.2/M, score 20.2
OpenAI: GPT-4.1 Nano (batch) — $0.05/M, score 11.1
GPT-4.1 nano — $0.1/M, score 11.1
GPT OSS 20B — $0.02/M, score 20.7
GPT-5.4 mini — $0.75/M, score 56.1
GPT-4o mini — $0.15/M, score 11.4
Trinity Large Thinking — $0.25/M, score 25.8
Qwen: Qwen3 30B A3B Thinking 2507 — $0.2/M, score 12.1
Z.ai: GLM 5.2 — $0.966/M, score 68.8
Qwen: Qwen3.7 Max — $1.475/M, score 66.0
SpaceXAI: Grok Build 0.1 — $1/M, score 51.5
Qwen: Qwen3.6 35B A3B — $0.1/M, score 41.9
Qwen: Qwen3.6 27B — $0.3/M, score 53.7
inclusionAI: Ling-2.6-flash — $0.01/M, score 25.3
Mistral: Mistral Small 4 — $0.15/M, score 26.6
Kwaipilot: KAT-Coder-Pro V2 — $0.3/M, score 59.5
Qwen: Qwen3.5-9B — $0.1/M, score 28.7
Qwen: Qwen3.5-35B-A3B — $0.312/M, score 37.0
Qwen: Qwen3.5-122B-A10B — $0.26/M, score 45.7
Upstage: Solar Pro 3 — $0.15/M, score 16.2
Z.ai: GLM 4.7 — $0.4/M, score 45.3
Amazon: Nova 2 Lite — $0.3/M, score 23.0
Mistral: Ministral 3 3B 2512 — $0.1/M, score 4.8
Google: Gemma 3n 4B — $0.06/M, score 3.2
Meta: Llama 4 Maverick — $0.2/M, score 16.3
OpenAI: o3 Mini High — $1.1/M, score 16.3
Meta: Llama 3.3 70B Instruct — $0.1/M, score 11.9
Meta: Llama 3.1 8B Instruct — $0.05/M, score 5.4
Nex AGI: Nex-N2-Pro — $0.25/M, score 59.1
Mistral: Mistral Medium 3.5 — $1.5/M, score 46.9
inclusionAI: Ring-2.6-1T — $0.075/M, score 42.8
IBM: Granite 4.1 8B — $0.05/M, score 9.5
Qwen: Qwen3.5 397B A17B — $0.55/M, score 48.2
Qwen: Qwen3 Coder Next — $0.12/M, score 36.2
Mistral: Ministral 3 14B 2512 — $0.2/M, score 14.4
Anthropic: Claude Haiku 4.5 — $1/M, score 43.9
Qwen: Qwen3 235B A22B Thinking 2507 — $0.23/M, score 22.1
Qwen: Qwen3 8B — $0.117/M, score 9.0
Qwen: Qwen3 14B — $0.227/M, score 13.8
Qwen: Qwen3 32B — $0.08/M, score 15.3
Meta: Llama 4 Scout — $0.1/M, score 8.2
DeepSeek: DeepSeek V3 0324 — $0.25/M, score 21.2
Cohere: Command A — $2.5/M, score 27.8
Step 3.7 Flash — $0.185/M, score 39.6
Gemini 3.5 Flash Lite
Input price per million tokens (log scale)
Index
The stepped line is the efficient frontier: at each price, the best score available for that money or less. A model sitting on it is not being beaten by anything cheaper. Price is log-scaled because listed rates span four orders of magnitude. Only models with both a listed price and a score on this benchmark can appear.
Catalog activity
Change log Field-level changes detected between successful source imports.
Full change log
No changes recorded This record has not changed within the retained import history.
Provenance
Sources & verification Every figure on this page traces back to one of these records.
Methodology
Last verified Sep 11, 2026
Status Source-linked
Confidence Medium
Catalog source Models.dev
Public API
Use this record Fetch the complete source-linked model record. No key, no account, no rate-limited tier.
API documentation
Endpoint Copy
GET https://model.kyssta.lol/api/v1/models/google/gemini-3.5-flash-lite
curl Copy
curl "https://model.kyssta.lol/api/v1/models/google/gemini-3.5-flash-lite"
Common questions
Frequently asked questions Answered directly from the stored record — nothing here is generated beyond the catalog's own fields.
What is Gemini 3.5 Flash Lite?
Fast Gemini model balancing multimodal reasoning, tool use, and cost. It is published by Google and catalogued here from Models.dev.
How much does Gemini 3.5 Flash Lite cost?
Listed input pricing starts at $0.15 per million tokens from Kilo Gateway, rising to $0.33 across 23 listed providers.
What is the context length of Gemini 3.5 Flash Lite?
Gemini 3.5 Flash Lite accepts up to 1.04858M tokens of context and returns up to 65.536K output tokens.
Does Gemini 3.5 Flash Lite support tool calling and structured output?
Provider catalogs list support for tool calling, structured output, reasoning, and image input.
Which providers serve Gemini 3.5 Flash Lite?
23 providers list this model: SAP AI Core, Kilo Gateway, Eden AI, Impossibl, Google, NanoGPT and 17 more.
Are the weights for Gemini 3.5 Flash Lite open?
No. This model is served through hosted APIs only.
When was Gemini 3.5 Flash Lite released?
The catalog records a release date of 2026-07-21, last verified Sep 11, 2026.