Provider guide
Where to run Gemini 3.8 Flash
8 live listings tracked.
FASTEST MEASURED
152 tok/s
measured throughput · $1.35 input / $6.75 output per million tokens
Google Vertex AIzero-retention$0.75 in/1M$3.75 out/1M64 tok/s1M—2 hours agoOpenRouter$0.75 in/1M$3.75 out/1M—1M—2 hours agoGoogle AI Studio$0.75 in/1M$3.75 out/1M129 tok/s1M—2 hours agoGoogle AICHEAPEST$0.75 in/1M$3.75 out/1M—1M—2 hours agoGoogle AI Studioflex tier$0.38 in/1M$1.88 out/1M144 tok/s1M—2 hours agoGoogle Vertex AIflex tierzero-retention$0.38 in/1M$1.88 out/1M25 tok/s1M—2 hours agoGoogle Vertex AIpriority tierzero-retention$1.35 in/1M$6.75 out/1M101 tok/s1M—2 hours agoGoogle AI Studiopriority tier$1.35 in/1M$6.75 out/1M152 tok/s1M—2 hours ago
Full specs, hardware verdicts and benchmarks on the Gemini 3.8 Flash model page →