Models / Google/ Gemini 3.5 Flash Lite

Gemini 3.5 Flash Lite

Google · released Jul 21, 2026

Input: text, images, audio, video and documents. Output: text.InputOutput
Type
Closed
Input
$0.30
Output
$2.50
Cached
None held

List price · per 1M tokens · Google AI at 1M context · machine-readable source ↗

Our take

Written Sep 2, 2026

Gemini 3.5 Flash Lite is a hosted-only multimodal model from Google that accepts text, images, files, audio and video, and can handle up to one million tokens in a single request. Released in July 2026, it is positioned as the lighter sibling to the full Flash line, with measured strengths in coding and a very large context window for its price class.

Who should pick it

Pick this for long-context document and media analysis where you need broad input support, or for budget-conscious coding tasks where its measured coding scores lead its other skills. Use it when proprietary hosting is acceptable and you can navigate tier selection carefully. Skip it if you need agentic or autonomous task performance, predictable throughput, or if you want to self-host.

The case for it

  • One-million-token request limit, extremely large for its price class.
  • Coding is the standout skill: LiveBench Coding 76.07%, well above its overall score, and Arena Coding 1499.86 leads its own profile by 43 points.
  • Lowest tier is half the price of the next one up on the same model.

The case against it

  • Agentic performance is weak across every measured dimension, with negative scores on task outcome, recovery and steerability.
  • LiveBench Data Analysis 53.25% and Agentic Coding 45.25% are particular soft spots, the latter nearly 19 points below its overall score.
  • Throughput varies wildly from 17 to 108 tokens per second depending on tier, with no way to predict which you will get.
00

How good is it?

A closed text model for everyday questions and drafting prose, though it struggles with multi-step agent work and changing course.

Good at
  • getting answers to everyday questionsArena Text (overall) · 42nd of 168
  • drafts, rewrites and editingArena Creative Writing · 39th of 168
Less good at
  • multi-step work it carries out for youArena Agent · 52nd of 55
  • calling tools to carry out requestsArena Agent · Tool use · 49th of 55
  • changing course when you give new instructionsArena Agent · Steerability · 55th of 55
  • getting back on track after a step failsArena Agent · Recovery · 51st of 55

EverydayGeneral questions and everyday reasoning

3.5 of 5

Arena Text (overall)42nd of 168 · 1456

Arena Hard Prompts 47th of 168Arena Maths 55th of 163LiveBench Data Analysis 58th of 58LiveBench Mathematics 58th of 58LiveBench Reasoning 58th of 58

CodingWriting and fixing code on its own

4 of 5

Arena Coding46th of 168 · 1503

Arena Code (WebDev) 49th of 95LiveBench Coding 37th of 58

AgenticPlanning, calling tools, staying on task

1 of 5

Arena Agent52nd of 55 · −0.153

LiveBench Agentic Coding 46th of 58

WritingDrafting and rewriting prose

3 of 5

Arena Creative Writing39th of 168 · 1435

LiveBench Language 52nd of 58
How it behaves in an agent loop
Tool usereaches for the right one, and does not invent one49th of 55
Steerabilitydoes what it was asked, and changes course when told55th of 55
Recoverygets back on track after a command fails51st of 55
Task outcomefinishes what the session set out to do51st of 55

Placings on Arena's agent boards, from live sessions people ran themselves. A model can lead on one of these and sit mid-field on the others.

Other boards it appears on
Arena Instruction Following 48th of 168LiveBench Instruction Following 31st of 58LiveBench 57th of 58Arena Agent · Tool use 49th of 55Arena Agent · Recovery 51st of 55Arena Agent · Task outcome 51st of 55Arena Agent · Steerability 55th of 55

Boards this model appears on that none of the ratings above are built on.

Every published score for this model20 scoresEvery figure we hold, from 20 boards, with who ran it and a link to the source — including the boards no rating above is built on.
LiveBenchreasoning
63.94source ↗
45.25source ↗
76.07source ↗
53.25source ↗
67.24source ↗
71.82source ↗
73.74source ↗
60.19source ↗
−0.153source ↗
−0.287source ↗
−0.128source ↗
−0.176source ↗
−0.015source ↗
1503source ↗
1435source ↗
1475source ↗
1440source ↗
1456source ↗
1439source ↗
01

Where to rent it

Prices checked 3 hours ago — each listing carries its own date.

Cheapest published offer

Google AI, direct

The lab is the cheapest at this context. The strip above and this offer are the same one, compared at 1M of context. 2 cheaper rows below are outside that comparison: a non-standard pricing tier.

per 1M tokens
$0.30 in / $2.50 out
Context served
1M
Throughput
Not measured
Current provider offers with price, context and prompt-privacy answers
ProviderIn / out per 1M tokensContextThroughputTrains on promptsLogs promptsZero retention
Google AI Studioflex tierThrough OpenRouter$0.15 / $1.25checked 3 hours ago1M66K max reply23 tok/sNoYes55 daysUnknown
Google Vertex AIflex tierglobalThrough OpenRouter$0.15 / $1.25checked 3 hours ago1M66K max reply20 tok/sNoNoConfirmed
Google Vertex AIglobalThrough OpenRouter$0.30 / $2.50checked 3 hours ago1M66K max reply66 tok/sNoNoConfirmed
Google AIDirect$0.30 / $2.50checked 3 hours ago1M66K max replynot measuredUnknownUnknownUnknown
Google AI StudioThrough OpenRouter$0.30 / $2.50checked 3 hours ago1M66K max reply87 tok/sNoYes55 daysUnknown
OpenRouterOpenRouter's own listing$0.30 / $2.50checked 3 hours ago1Mnot measuredUnknownUnknownUnknown
Google Vertex AIusThrough OpenRouter$0.33 / $2.75checked 3 hours ago1M66K max reply117 tok/sNoNoConfirmed
Google Vertex AIeuThrough OpenRouter$0.33 / $2.75checked 3 hours ago1M66K max reply70 tok/sNoNoConfirmed
Google AI Studiopriority tierThrough OpenRouter$0.54 / $4.50checked 3 hours ago1M66K max reply63 tok/sNoYes55 daysUnknown
Google Vertex AIpriority tierglobalThrough OpenRouter$0.54 / $4.50checked 3 hours ago1M66K max reply27 tok/sNoNoConfirmed

Across the 10 listings we hold: 8 say they do not train on prompts, 0 say they do and 2 do not say. 5 appear in the zero-retention registry we check; the rest are unknown to us.

What each host's API supports

From the parameter list each endpoint publishes. Streaming is omitted: nothing we hold reports it, for any model.

API features per host
ProviderTool callingJSON outputStrict schema
Google AI StudioflexThrough OpenRouter✓✓✓
Google Vertex AIflex · globalThrough OpenRouter✓✓✓
Google Vertex AIglobalThrough OpenRouter✓✓✓
Google AIDirect
Google AI StudioThrough OpenRouter✓✓✓
OpenRouterOpenRouter's own listing✓✓✓
Google Vertex AIusThrough OpenRouter✓✓✓
Google Vertex AIeuThrough OpenRouter✓✓✓
Google AI StudiopriorityThrough OpenRouter✓✓✓
Google Vertex AIpriority · globalThrough OpenRouter✓✓✓

Tool calling: 9 of 10 listings say yes, 1 publishes no parameter list. JSON output: 9 of 10 listings say yes, 1 publishes no parameter list. Strict schema: 9 of 10 listings say yes, 1 publishes no parameter list.

02

Models people weigh against Gemini 3.5 Flash Lite

03

When we formed this view

Recent changes

Sep 25, 2026BenchmarkScored 1503 on Arena Coding
What movedleaderboard
Sep 25, 2026BenchmarkScored 1435 on Arena Creative Writing
What movedleaderboard
Sep 25, 2026BenchmarkScored 1475 on Arena Hard Prompts
What movedleaderboard
Sep 25, 2026BenchmarkScored 1444 on Arena Instruction Following
What movedleaderboard
Sep 25, 2026BenchmarkScored 1440 on Arena Maths
What movedleaderboard
Sep 25, 2026BenchmarkScored 1456 on Arena Text (overall)
What movedleaderboard
Sep 25, 2026BenchmarkScored 1439 on Arena Code (WebDev)
What movedleaderboard
Sep 15, 2026BenchmarkScored −0.153 on Arena Agent
What movedleaderboard
Sep 15, 2026BenchmarkScored −0.287 on Arena Agent · Recovery
What movedleaderboard
Sep 15, 2026BenchmarkScored −0.128 on Arena Agent · Steerability
What movedleaderboard

Each date is the day we first saw the change, or the day the maker announced it.

What we do not know about this model yet

  • 1 of 10 listings publishes no parameter list, so what its API accepts is unknown to us.
  • Nothing we hold says whether an endpoint streams, so we do not show it either way.
  • 2 of 10 listings do not say whether they train on prompts.
  • We hold no batch or off-peak rate for any of its listings.
04

Licence and identifiers

What the licence allowsWe hold no licence record for this model. Inside are the identifiers you need to pull it — its Hugging Face repo where we have one, our slug and a machine-readable card.

Licence

We hold no licence record for this model, and no record of published weights either — so we can neither summarise its terms nor point you at the weights.

Identifiers

Takes in, gives back
Text, images, audio, video and documents in, text out
Catalogue slug
google-gemini-3-5-flash-lite

Machine-readable model card (omc.json) →

Something wrong on this page? Tell us