Models /GLM 4.7 /Where to run
Provider guide

Where to run GLM 4.7

10 live offers tracked — output prices vary 1.9× between the cheapest and most expensive host, so the provider choice matters as much as the model choice.

CHEAPEST
$0.40 / $1.75
per 1M tokens in / out
FASTEST MEASURED
225 tok/s
measured throughput · $2.25 input / $2.75 output per million tokens
OpenRouterCHEAPEST$0.40 in/1M$1.75 out/1M205K14m agoDeepInfra$0.40 in/1M$1.75 out/1M203Kfp46h agoAtlasCloud$0.52 in/1M$1.85 out/1M37 tok/s203Kfp86h agoNovita AI$0.60 in/1M$2.20 out/1M205K6h agoPhala$0.85 in/1M$3.30 out/1M33 tok/s131K5d agoDeepInfrazero-retention$0.40 in/1M$1.75 out/1M30 tok/s203Kfp412m agoNovita AIzero-retention$0.54 in/1M$1.98 out/1M28 tok/s205Kfp86h agoGoogle Vertex AIzero-retention$0.60 in/1M$2.20 out/1M111 tok/s200K6h agoVenice AIzero-retention$0.55 in/1M$2.65 out/1M26 tok/s198Kfp412m agoCerebraszero-retention$2.25 in/1M$2.75 out/1M225 tok/s131Kfp1612m ago
Full specs, hardware verdicts and benchmarks on theGLM 4.7 model page →