Provider guide
Where to run Gemini 3.1 Flash Lite
11 live listings tracked.
FASTEST MEASURED
129 tok/s
measured throughput · $0.28 input / $1.65 output per million tokens
Google AICHEAPEST$0.25 in/1M$1.50 out/1M—1M—9 hours agoGoogle Vertex AIzero-retention$0.25 in/1M$1.50 out/1M66 tok/s1M—3 hours agoOpenRouter$0.25 in/1M$1.50 out/1M—1M—3 hours agoDeepInfra$0.25 in/1M$1.50 out/1M—1M—9 hours agoGoogle AI Studio$0.25 in/1M$1.50 out/1M93 tok/s1M—3 hours agoGoogle Vertex AIzero-retention$0.28 in/1M$1.65 out/1M129 tok/s1M—3 hours agoGoogle Vertex AIzero-retention$0.28 in/1M$1.65 out/1M17 tok/s1M—3 hours agoGoogle AI Studioflex tier$0.13 in/1M$0.75 out/1M51 tok/s1M—3 hours agoGoogle Vertex AIflex tierzero-retention$0.13 in/1M$0.75 out/1M2 tok/s1M—3 hours agoGoogle AI Studiopriority tier$0.45 in/1M$2.70 out/1M85 tok/s1M—3 hours agoGoogle Vertex AIpriority tierzero-retention$0.45 in/1M$2.70 out/1M101 tok/s1M—3 hours ago
Full specs, hardware verdicts and benchmarks on the Gemini 3.1 Flash Lite model page →