Provider guide
Where to run Gemini 2.5 Flash
10 live listings tracked.
FASTEST MEASURED
194 tok/s
measured throughput · $0.15 input / $1.25 output per million tokens
DeepInfra$0.30 in/1M$2.50 out/1M—1M—2 hours agoGoogle AICHEAPEST$0.30 in/1M$2.50 out/1M—1M—2 hours agoGoogle Vertex AIzero-retention$0.30 in/1M$2.50 out/1M97 tok/s1M—2 hours agoGoogle Vertex AIzero-retention$0.30 in/1M$2.50 out/1M56 tok/s1M—14 hours agoGoogle Vertex AIzero-retention$0.30 in/1M$2.50 out/1M73 tok/s1M—2 hours agoOpenRouter$0.30 in/1M$2.50 out/1M—1M—2 hours agoGoogle AI Studio$0.30 in/1M$2.50 out/1M49 tok/s1M—2 hours agoGoogle AI Studioflex tier$0.15 in/1M$1.25 out/1M194 tok/s1M—2 hours agoGoogle AI Studiopriority tier$0.54 in/1M$4.50 out/1M110 tok/s1M—2 hours agoGoogle Vertex AIpriority tierzero-retention$0.54 in/1M$4.50 out/1M38 tok/s1M—2 hours ago
Full specs, hardware verdicts and benchmarks on the Gemini 2.5 Flash model page →