Models /Qwen3.5 397B A17B /Where to run
Provider guide

Where to run Qwen3.5 397B A17B

11 live listings tracked — output prices vary 1.9× between the cheapest and most expensive host, so the provider choice matters as much as the model choice.

CHEAPEST
$0.39 / $2.34
per 1M tokens in / out
FASTEST MEASURED
131 tok/s
measured throughput · $0.60 input / $3.60 output per million tokens
Alibaba CloudCHEAPEST$0.39 in/1M$2.34 out/1M71 tok/s262K—2 hours agoDeepInfrazero-retention through OpenRouterDirect and through OpenRouter$0.45 in/1M$3.00 out/1M38 tok/sthrough OpenRouter262Kfp82 hours agoOpenRouter$0.55 in/1M$3.50 out/1M—262K—2 hours agoDigitalOcean Gradientzero-retention$0.55 in/1M$3.50 out/1M2 tok/s131K—2 hours agoAtlasCloud$0.55 in/1M$3.50 out/1M74 tok/s262Kfp82 hours agoPhalazero-retention$0.55 in/1M$3.50 out/1M44 tok/s262K—2 hours agoStreamLake$0.60 in/1M$3.60 out/1M131 tok/s256K—2 hours agoNovita AIzero-retention through OpenRouterDirect and through OpenRouter$0.60 in/1M$3.60 out/1M63 tok/sthrough OpenRouter262K—2 hours agoGMICloud$0.60 in/1M$3.60 out/1M74 tok/s262Kfp82 hours agoParasailzero-retention$0.50 in/1M$3.60 out/1M43 tok/s262Kfp82 hours agoVenice AIzero-retention$0.75 in/1M$4.50 out/1M31 tok/s128K—2 hours ago
Full specs, hardware verdicts and benchmarks on the Qwen3.5 397B A17B model page →