Gemini 3.5 Flash Lite vs GPT-5.6 Luna vs Laguna S 2.1
Compare Gemini 3.5 Flash Lite and GPT-5.6 Luna and Laguna S 2.1 on key metrics including pricing, context length, capabilities, providers, and published benchmarks.
Overview
Author
Google
Context length
1.04858Mtokens
Reasoning
Supported
Input modalities
Output modalities
Providers
24 providers
Author
OpenAI
Context length
1.05MtokensBest
Reasoning
Supported
Input modalities
Output modalities
Providers
37 providers
Author
Poolside
Context length
1.04858Mtokens
Reasoning
Supported
Input modalities
Output modalities
Providers
7 providers
Pricing per million tokens
Input
$0.3
Output
$2.5
Cached input
$0.03
Cache write
Not listed
Input
$0.2
Output
$1.2
Cached input
$0.02
Cache write
$0.25
Input
$0.09Best
Output
$0.18Best
Cached input
$0.009
Cache write
Not listed
Limits
Context length
1.04858Mtokens
Max output
65.536Ktokens
Released
2026-07-21
Knowledge cutoff
2026-03
Context length
1.05MtokensBest
Max output
128KtokensBest
Released
2026-07-09
Knowledge cutoff
2026-02-16
Context length
1.04858Mtokens
Max output
32.768Ktokens
Released
2026-07-21
Knowledge cutoff
Not documented
Capabilities
Tool calling
Supported
Structured output
Supported
Vision input
Supported
Attachments
Supported
Tool calling
Supported
Structured output
Supported
Vision input
Supported
Attachments
Supported
Tool calling
Supported
Structured output
Not supported
Vision input
Not supported
Attachments
Not supported
Availability
Inference providers
24providers
Provider list
302aiGoogleabacuscortecscrossmodel+17 more
Inference providers
37providersBest
Provider list
302aiOpenAIabacusai-routeraihubmix+31 more
Inference providers
7providers
Provider list
kilonano-gptopenrouterpioneerpoolside+1 more
Benchmarks
- Gemini 3.5 Flash Lite
- GPT-5.6 Luna
* Laguna S 2.1 has no intelligence data
- Gemini 3.5 Flash Lite
- GPT-5.6 Luna
* Laguna S 2.1 has no coding data
- Gemini 3.5 Flash Lite
- GPT-5.6 Luna
* Laguna S 2.1 has no agentic data
All published benchmark results 21 variants
Agentic Index
index15.9
Agents' Last Exam
scoreNot reported
Artificial Analysis Coding Agent Index
index score · 1.1 · CodexNot reported
Artificial Analysis Intelligence Index
index score · 4.1Not reported
BrowseComp
accuracyNot reported
CharXiv Reasoning
accuracy76.5Best
Coding Index
index49.3
DeepSWE
resolve rate · 1.1Not reported
FrontierMath
accuracy · v2Not reported
GDM-MRCR
accuracy · v221.3Best
GDPval-AA
Elo · v21140.0Best
GPQA Diamond
accuracyNot reported
Intelligence Index
index22.7
MLE-Bench
average position score39.2Best
MMMU Pro
accuracyNot reported
OSWorld
success rate · 2.0Not reported
OSWorld-Verified
success rate74.0Best
SWE-Bench Pro
resolve rate54.2
Terminal-Bench
success rate · 2.1Not reported
Terminal-Bench
accuracy · 2.1 · Terminus 254.0Best
Toolathlon
success rateNot reported
Agentic Index
index42.7Best
Agents' Last Exam
score50.3Best
Artificial Analysis Coding Agent Index
index score · 1.1 · Codex74.6Best
Artificial Analysis Intelligence Index
index score · 4.151.2Best
BrowseComp
accuracy83.3Best
CharXiv Reasoning
accuracyNot reported
Coding Index
index71.4Best
DeepSWE
resolve rate · 1.167.2Best
FrontierMath
accuracy · v278.6Best
GDM-MRCR
accuracy · v2Not reported
GDPval-AA
Elo · v2Not reported
GPQA Diamond
accuracy92.3Best
Intelligence Index
index37.5Best
MLE-Bench
average position scoreNot reported
MMMU Pro
accuracy78.4Best
OSWorld
success rate · 2.045.6Best
OSWorld-Verified
success rateNot reported
SWE-Bench Pro
resolve rate62.7Best
Terminal-Bench
success rate · 2.184.7Best
Terminal-Bench
accuracy · 2.1 · Terminus 2Not reported
Toolathlon
success rate53.4Best
Agentic Index
indexNot reported
Agents' Last Exam
scoreNot reported
Artificial Analysis Coding Agent Index
index score · 1.1 · CodexNot reported
Artificial Analysis Intelligence Index
index score · 4.1Not reported
BrowseComp
accuracyNot reported
CharXiv Reasoning
accuracyNot reported
Coding Index
indexNot reported
DeepSWE
resolve rate · 1.1Not reported
FrontierMath
accuracy · v2Not reported
GDM-MRCR
accuracy · v2Not reported
GDPval-AA
Elo · v2Not reported
GPQA Diamond
accuracyNot reported
Intelligence Index
indexNot reported
MLE-Bench
average position scoreNot reported
MMMU Pro
accuracyNot reported
OSWorld
success rate · 2.0Not reported
OSWorld-Verified
success rateNot reported
SWE-Bench Pro
resolve rateNot reported
Terminal-Bench
success rate · 2.1Not reported
Terminal-Bench
accuracy · 2.1 · Terminus 2Not reported
Toolathlon
success rateNot reported
Provenance
“Not reported” means the source catalog is silent on that field. Benchmark rows keep each published version and harness separate.