Models /GLM 5 /Where to run
Provider guide

Where to run GLM 5

10 live listings tracked — output prices vary 1.7× between the cheapest and most expensive host, so the provider choice matters as much as the model choice.

CHEAPEST
$0.60 / $1.92
per 1M tokens in / out
FASTEST MEASURED
61 tok/s
measured throughput · $0.60 input / $1.92 output per million tokens
OpenRouterCHEAPEST$0.60 in/1M$1.92 out/1M—205K—2 hours agoStreamLake$0.60 in/1M$1.92 out/1M59 tok/s198Kfp82 hours agoGMICloud$0.60 in/1M$1.92 out/1M61 tok/s203Kfp82 hours agoDeepInfra$0.60 in/1M$2.08 out/1M—203Kfp428 days agoBaidu$0.70 in/1M$2.24 out/1M50 tok/s203Kfp82 hours agoSiliconFlowzero-retention$0.95 in/1M$2.55 out/1M53 tok/s205Kfp82 hours agoVenice AIzero-retention$1.00 in/1M$3.20 out/1M57 tok/s198Kfp82 hours agoNovita AIzero-retention through OpenRouterDirect and through OpenRouter$1.00 in/1M$3.20 out/1M41 tok/sthrough OpenRouter203Kfp82 hours agoAmazon Bedrockzero-retention$1.00 in/1M$3.20 out/1M59 tok/s203K—2 hours agoZ.AIzero-retention$1.00 in/1M$3.20 out/1M50 tok/s203Kfp82 hours ago
Full specs, hardware verdicts and benchmarks on the GLM 5 model page →