Provider guide
Where to run GLM 5.3
36 live listings tracked — output prices vary 3.9× between the cheapest and most expensive host, so the provider choice matters as much as the model choice.
Rekazero-retention$0.37 in/1M$1.14 out/1M60 tok/s262K—9 hours agoSiliconFlowzero-retention$0.70 in/1M$2.20 out/1M54 tok/s1Mfp83 hours agoInferenceNetzero-retentionCHEAPEST$0.68 in/1M$2.28 out/1M91 tok/s1M—2 days agoNovita AIzero-retention through OpenRouterDirect and through OpenRouter$1.40 in/1M$0.78 in/1M$4.40 out/1Mdirect$2.46 out/1Mthrough OpenRouter45 tok/sthrough OpenRouter1Mfp83 hours agoDeepInfrazero-retention through OpenRouterDirect and through OpenRouter$0.90 in/1M$0.56 in/1M$4.00 out/1Mdirect$2.50 out/1Mthrough OpenRouter36 tok/sthrough OpenRouter1Mfp43 hours agoPhalazero-retention$0.84 in/1M$2.64 out/1M44 tok/s1M—3 hours agoDigitalOcean Gradientzero-retention$0.91 in/1M$2.86 out/1M47 tok/s1M—3 hours agoMorphzero-retention$1.19 in/1M$2.99 out/1M109 tok/s1Mfp82 days agoGMICloud$0.98 in/1M$3.08 out/1M53 tok/s1Mfp83 hours agoInceptronzero-retention$0.60 in/1M$3.39 out/1M60 tok/s1Mfp43 hours agoAkashMLzero-retention$1.05 in/1M$3.56 out/1M80 tok/s1Mfp83 hours agoDecartzero-retention$1.19 in/1M$3.74 out/1M129 tok/s1Mfp43 hours agoAlibaba Cloud$1.19 in/1M$3.74 out/1M74 tok/s1M—3 hours agoMakorazero-retention$0.85 in/1M$3.93 out/1M109 tok/s980Kfp49 hours agoFriendli$1.26 in/1M$3.96 out/1M92 tok/s1M—3 hours agoSail Research$0.77 in/1M$4.00 out/1M83 tok/s1Mfp83 days agoSail Researchzero-retention$0.77 in/1M$4.00 out/1M74 tok/s1Mfp83 days agoRelacezero-retention$0.15 in/1M$4.00 out/1M103 tok/s1M—3 hours agoOpenRouter$1.40 in/1M$4.40 out/1M—1M—26 hours agoMistral AIzero-retention$1.40 in/1M$4.40 out/1M127 tok/s1Mnvfp43 hours agoTogether AIzero-retention$1.40 in/1M$4.40 out/1M127 tok/s1M—3 hours agoVenice AIzero-retention$1.40 in/1M$4.40 out/1M43 tok/s1M—3 hours agoCloudflare Workers AI$1.40 in/1M$4.40 out/1M33 tok/s1M—3 hours agoAtlasCloud$1.40 in/1M$4.40 out/1M50 tok/s1Mfp83 hours agoCrusoezero-retention$1.40 in/1M$4.40 out/1M114 tok/s1Mfp43 hours agoParasailzero-retention$1.40 in/1M$4.40 out/1M110 tok/s1Mfp83 hours agoBasetenzero-retention$1.40 in/1M$4.40 out/1M65 tok/s1Mfp43 hours agoBaidu$1.40 in/1M$4.40 out/1M83 tok/s1Mfp821 hours agoWaferzero-retention$1.40 in/1M$4.40 out/1M77 tok/s1M—3 days agoWaferzero-retention$1.82 in/1M$4.40 out/1M95 tok/s1M—4 days agoFireworks AIzero-retention$1.40 in/1M$4.40 out/1M43 tok/s1M—3 hours agoZ.AIzero-retention$1.40 in/1M$4.40 out/1M57 tok/s1Mfp83 hours agoModalzero-retention$1.40 in/1M$4.40 out/1M62 tok/s1M—3 hours agoPrimeIntellectzero-retention$1.40 in/1M$4.40 out/1M127 tok/s1M—3 hours agoBasetenfast tier$2.10 in/1M$6.60 out/1M141 tok/s1Mfp89 hours agoAlibaba Cloudfast tier$2.80 in/1M$8.80 out/1M73 tok/s1M—3 hours ago
Full specs, hardware verdicts and benchmarks on the GLM 5.3 model page →