Models /Qwen3 VL 235B A22B Instruct /Where to run
Provider guide

Where to run Qwen3 VL 235B A22B Instruct

9 live offers tracked — output prices vary 2.2× between the cheapest and most expensive host, so the provider choice matters as much as the model choice.

CHEAPEST
$0.21 / $1.90
per 1M tokens in / out
FASTEST MEASURED
36 tok/s
measured throughput · $0.21 input / $1.90 output per million tokens
DeepInfra$0.20 in/1M$0.88 out/1M262Kfp86h agoAlibaba Cloud$0.26 in/1M$1.04 out/1M24 tok/s131K6d agoAlibaba Cloud$0.26 in/1M$1.04 out/1M35 tok/s131Kfp815m agoNovita AI$0.30 in/1M$1.50 out/1M131K6h agoOpenRouterCHEAPEST$0.21 in/1M$1.90 out/1M262K17m agoDeepInfrazero-retention$0.20 in/1M$0.88 out/1M21 tok/s262Kfp814h agoNovita AIzero-retention$0.30 in/1M$1.50 out/1M16 tok/s131Kbf1644h agoParasailzero-retention$0.21 in/1M$1.90 out/1M36 tok/s131Kfp815m agoVenice AIzero-retention$0.21 in/1M$1.90 out/1M9 tok/s128Kfp820h ago
Full specs, hardware verdicts and benchmarks on theQwen3 VL 235B A22B Instruct model page →