Gemini 3.7 Flash
Google · released Aug 13, 2026
- Type
- Closed
- Input
- $0.75
- Output
- $3.75
- Cached
- None held
List price · per 1M tokens · Google AI at 1M context · machine-readable source ↗
Our take
Written Sep 14, 2026Gemini 3.7 Flash is Google's multimodal workhorse that handles up to one million tokens in a single request across text, images, audio, video and files. It excels at mathematics and web-development tasks, though its agentic behaviour scores are weak and throughput varies sharply depending on which Google endpoint you choose.
Pick this for long-document or rich-media workflows that need a million-token context, or for mathematics and web-development tasks where its measured scores are strongest. Use it if you are already in the Google Cloud ecosystem and want the cheapest tier through Vertex or AI Studio. Skip it if you need reliable agentic steerability, strong data-analysis performance, or predictable throughput across endpoints.
The case for it
- Exceptional mathematics performance on refreshed competition problems, at 93.47% on LiveBench Mathematics.
- Highest subscore in human preference voting for web-development tasks, at 1587.3 on Arena Code (WebDev).
- One-million-token request limit with multimodal input spanning images, audio, video and files.
- Same lowest price available through two Google channels, creating genuine optionality.
The case against it
- Weak agentic behaviour scores: steerability and recovery are both negative, with only task outcome positive.
- Data analysis and agentic coding lag well behind its own strong reasoning scores, by over 20 points each.
- Throughput varies more than twofold across endpoints, and the same provider charges several times more across tiers.
How good is it?
A general-purpose text model for everyday questions, drafting and coding, with no measured weak spot on this board.
- answering everyday questionsArena Text (overall) · 11th of 168
- drafting and editing proseArena Creative Writing · 3rd of 168
- writing and completing codeArena Coding · 27th of 168
EverydayGeneral questions and everyday reasoning
Arena Text (overall)11th of 168 · 1488
CodingWriting and fixing code on its own
Arena Coding27th of 168 · 1519
AgenticPlanning, calling tools, staying on task
Arena Agent32nd of 55 · −0.012
WritingDrafting and rewriting prose
Arena Creative Writing3rd of 168 · 1492
Placings on Arena's agent boards, from live sessions people ran themselves. A model can lead on one of these and sit mid-field on the others.
Boards this model appears on that none of the ratings above are built on.
Every published score for this model20 scoresEvery figure we hold, from 20 boards, with who ran it and a link to the source — including the boards no rating above is built on.
Where to rent it
Prices checked 3 hours ago — each listing carries its own date.
Google AI, direct
The lab is the cheapest at this context. The strip above and this offer are the same one, compared at 1M of context. 2 cheaper rows below are outside that comparison: a non-standard pricing tier.
- per 1M tokens
- $0.75 in / $3.75 out
- Context served
- 1M
- Throughput
- Not measured
| Provider | In / out per 1M tokens | Context | Throughput | Trains on prompts | Logs prompts | Zero retention |
|---|---|---|---|---|---|---|
| Google AI Studioflex tierThrough OpenRouter | $0.38 / $1.88checked 3 hours ago | 1M66K max reply | 139 tok/s | No | Yes55 days | Unknown |
| Google Vertex AIflex tierglobalThrough OpenRouter | $0.38 / $1.88checked 3 hours ago | 1M66K max reply | 20 tok/s | No | No | Confirmed |
| Google AIDirect | $0.75 / $3.75checked 3 hours ago | 1M66K max reply | not measured | Unknown | Unknown | Unknown |
| Google Vertex AIglobalThrough OpenRouter | $0.75 / $3.75checked 3 hours ago | 1M66K max reply | 72 tok/s | No | No | Confirmed |
| Google AI StudioThrough OpenRouter | $0.75 / $3.75checked 3 hours ago | 1M66K max reply | 128 tok/s | No | Yes55 days | Unknown |
| DeepInfraDirect | $0.75 / $3.75checked 3 hours ago | 1M | not measured | Unknown | Unknown | Unknown |
| OpenRouterOpenRouter's own listing | $0.75 / $3.75checked 3 hours ago | 1M | not measured | Unknown | Unknown | Unknown |
| Google AI Studiopriority tierThrough OpenRouter | $1.35 / $6.75checked 3 hours ago | 1M66K max reply | 87 tok/s | No | Yes55 days | Unknown |
| Google Vertex AIpriority tierglobalThrough OpenRouter | $1.35 / $6.75checked 3 hours ago | 1M66K max reply | 106 tok/s | No | No | Confirmed |
Across the 9 listings we hold: 6 say they do not train on prompts, 0 say they do and 3 do not say. 3 appear in the zero-retention registry we check; the rest are unknown to us.
What each host's API supports
From the parameter list each endpoint publishes. Streaming is omitted: nothing we hold reports it, for any model.
| Provider | Tool calling | JSON output | Strict schema |
|---|---|---|---|
| Google AI StudioflexThrough OpenRouter | ✓ | ✓ | ✓ |
| Google Vertex AIflex · globalThrough OpenRouter | ✓ | ✓ | ✓ |
| Google AIDirect | |||
| Google Vertex AIglobalThrough OpenRouter | ✓ | ✓ | ✓ |
| Google AI StudioThrough OpenRouter | ✓ | ✓ | ✓ |
| DeepInfraDirect | |||
| OpenRouterOpenRouter's own listing | ✓ | ✓ | ✓ |
| Google AI StudiopriorityThrough OpenRouter | ✓ | ✓ | ✓ |
| Google Vertex AIpriority · globalThrough OpenRouter | ✓ | ✓ | ✓ |
Tool calling: 7 of 9 listings say yes, 2 publish no parameter list. JSON output: 7 of 9 listings say yes, 2 publish no parameter list. Strict schema: 7 of 9 listings say yes, 2 publish no parameter list.
Models people weigh against Gemini 3.7 Flash
When we formed this view
Recent changes
Each date is the day we first saw the change, or the day the maker announced it.
What we do not know about this model yet
- 2 of 9 listings publish no parameter list, so what their API accepts is unknown to us.
- Nothing we hold says whether an endpoint streams, so we do not show it either way.
- 3 of 9 listings do not say whether they train on prompts.
- We hold no batch or off-peak rate for any of its listings.
Licence and identifiers
What the licence allowsWe hold no licence record for this model. Inside are the identifiers you need to pull it — its Hugging Face repo where we have one, our slug and a machine-readable card.
Licence
We hold no licence record for this model, and no record of published weights either — so we can neither summarise its terms nor point you at the weights.
Identifiers
- Takes in, gives back
- Text, images, audio, video and documents in, text out
- Catalogue slug
- google-gemini-3-7-flash