Models /Qwen3 32B /Where to run
Provider guide

Where to run Qwen3 32B

8 live offers tracked — output prices vary 1.6× between the cheapest and most expensive host, so the provider choice matters as much as the model choice.

CHEAPEST
$0.080 / $0.28
per 1M tokens in / out
FASTEST MEASURED
361 tok/s
measured throughput · $0.29 input / $0.59 output per million tokens
OpenRouterCHEAPEST$0.080 in/1M$0.28 out/1M131K15m agoDeepInfra$0.080 in/1M$0.28 out/1M41Kfp86h agoNebius AI Studio$0.10 in/1M$0.30 out/1M23 tok/s41Kfp86h agoNebius AI Studio$0.10 in/1M$0.30 out/1M25 tok/s41Kfp86d agoNovita AI$0.10 in/1M$0.45 out/1M41K6h agoDeepInfrazero-retention$0.080 in/1M$0.28 out/1M31 tok/s41Kfp814m agoSiliconFlowzero-retention$0.14 in/1M$0.57 out/1M16 tok/s131Kfp814m agoGroqzero-retention$0.29 in/1M$0.59 out/1M361 tok/s131K14m ago
Full specs, hardware verdicts and benchmarks on theQwen3 32B model page →