Models /GLM 4.6 /Where to run
Provider guide

Where to run GLM 4.6

5 live listings tracked.

CHEAPEST
$0.43 / $1.75
per 1M tokens in / out
FASTEST MEASURED
56 tok/s
measured throughput · $0.43 input / $1.75 output per million tokens
Venice AIzero-retention$0.43 in/1M$1.75 out/1M56 tok/s198Kfp42 hours agoOpenRouterCHEAPEST$0.43 in/1M$1.75 out/1M—205K—2 hours agoDeepInfrazero-retention through OpenRouterDirect and through OpenRouter$0.50 in/1M$2.00 out/1M20 tok/sthrough OpenRouter203Kfp42 hours agoNovita AIzero-retention through OpenRouterDirect and through OpenRouter$0.55 in/1M$2.20 out/1M25 tok/sthrough OpenRouter205Kbf162 hours agoZ.AIzero-retention$0.60 in/1M$2.20 out/1M25 tok/s203Kfp42 hours ago
Full specs, hardware verdicts and benchmarks on the GLM 4.6 model page →