Price from Google: the cheapest offer that is serving, at standard delivery and at the maker’s precision or better, at its standard rate.
Cheaper right now: $0.425/M at Google — delivery tier (flex) It does not pass the like-for-like test, so it never sets a rank.
Every price dimension · USD per million tokens
Input
$0.300/M
Output
$2.50/M
Cached input
$0.030/M
Cache write
$0.083/M
Cache write 1h
Unknown/M
Reasoning
$2.50/M
Web search
$0.014/call
Batch discount
50%
Reasoning tokens are billed separately at $2.50/M, on
top of output. Its share of your bill depends on the reasoning tokens used.
Who sells it
The same weights, different shops. Cheapest is not like-for-like when serving precision differs.
8 provider offers · rates, limits and conditions
Rates are USD per million tokens. Endpoint terms can differ even when the quoted price and precision match. A listed parameter is a provider declaration, not a task-success test.
Seller
Input
Output
Precision
Uptime (1d)
Seller
Googlegoogle-vertex/global/flex Endpoint terms
Cached input /M
$0.015
Context limit
1,048,576 tokens
Output limit
65,536 tokens
Tools
Listed by endpoint
Reasoning
Listed by endpoint
Promotional discount
None reported
Uptime · last 30 minutes
100.00%
Input$0.150
Output$1.25
Precisionundeclared
Uptime (1d)99.99%
Seller
Google AI Studiogoogle-ai-studio/flex Endpoint terms
Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.
Questions this page answers
What does Gemini 3.5 Flash Lite cost?
$0.300/M in, $2.50/M out in this catalogue, as of the fetch date on this page. That is the API row, not a subscription.
Does Gemini 3.5 Flash Lite have an independent quality score?
Gemini 3.5 Flash Lite has an LMArena score in this catalogue.
What beats Gemini 3.5 Flash Lite?
MiMo-V2.6-Pro has an equal-or-higher measured score and an equal-or-lower price, with capability trade-offs.
Does this page use Artificial Analysis scores for Gemini 3.5 Flash Lite?
No. Artificial Analysis figures are not published here. Quality on this page is LMArena Elo where a score exists; otherwise the row is unrated.
MODEL MONUMENT
Save or share this proof
Preview and more sharing options
Google: Gemini 3.5 Flash Lite. Balanced workload. Alternative has a trade-off. Evidence as of 06 Oct 2026.
Related
Higher score and cheaper, with a named tradeMiMo-V2.6-Pro
The verdict as one SVG, for a README or a docs page. It states this model's dominance
status at the balanced workload on the LMArena lens, with the date it was computed, and
it is rebuilt with the catalogue — so it changes when the verdict changes, including to
one you would rather it did not.
Markdown
[](https://undominated.ai/models/google__gemini-3.5-flash-lite/)