Models /DeepSeek V3.2 /Where to run
Provider guide

Where to run DeepSeek V3.2

19 live offers tracked — output prices vary 14.5× between the cheapest and most expensive host, so the provider choice matters as much as the model choice.

CHEAPEST
$0.27 / $0.40
per 1M tokens in / out
FASTEST MEASURED
40 tok/s
measured throughput · $0.37 input / $1.11 output per million tokens
Baidu$0.21 in/1M$0.31 out/1M26 tok/s131Kfp814m agoStreamLake$0.21 in/1M$0.32 out/1M20 tok/s128Kfp814m agoAtlasCloud$0.26 in/1M$0.38 out/1M21 tok/s164Kfp814m agoDeepInfra$0.26 in/1M$0.38 out/1M164Kfp46h agoNovita AICHEAPEST$0.27 in/1M$0.40 out/1M164K6h agoOpenRouter$0.27 in/1M$0.40 out/1M164K14m agoGMICloud$0.29 in/1M$0.43 out/1M28 tok/s164Kfp814m agoAlibaba Cloud$0.37 in/1M$1.11 out/1M40 tok/s131Kfp814m agoAlibaba Cloud$0.37 in/1M$1.11 out/1M34 tok/s131K6d agoFriendli$0.50 in/1M$1.50 out/1M26 tok/s164K14m agoSambaNova$3.00 in/1M$4.50 out/1M33K6h agoDeepInfrazero-retention$0.26 in/1M$0.38 out/1M12 tok/s164Kfp414m agoNovita AIzero-retention$0.27 in/1M$0.40 out/1M22 tok/s164Kfp814m agoSiliconFlowzero-retention$0.26 in/1M$0.42 out/1M25 tok/s164Kfp86h agoVenice AIzero-retention$0.33 in/1M$0.48 out/1M7 tok/s160K14m agoPhalazero-retention$1.00 in/1M$1.00 out/1M7 tok/s164K14m agoDigitalOcean Gradientzero-retention$0.42 in/1M$1.36 out/1M40 tok/s164K14m agoGoogle Vertex AIzero-retention$0.56 in/1M$1.68 out/1M14 tok/s164K14m agoSambaNovazero-retention$3.00 in/1M$4.50 out/1M38 tok/s33K14m ago
Full specs, hardware verdicts and benchmarks on theDeepSeek V3.2 model page →