Models /Qwen3 VL 235B A22B Thinking /Where to run
Provider guide

Where to run Qwen3 VL 235B A22B Thinking

5 live offers tracked — output prices vary 1.5× between the cheapest and most expensive host, so the provider choice matters as much as the model choice.

CHEAPEST
$0.26 / $2.60
per 1M tokens in / out
FASTEST MEASURED
60 tok/s
measured throughput · $0.40 input / $4.00 output per million tokens
Alibaba CloudCHEAPEST$0.26 in/1M$2.60 out/1M131K7d agoNovita AI$0.98 in/1M$3.95 out/1M131K6h agoOpenRouter$0.40 in/1M$4.00 out/1M131K32h agoAlibaba Cloud$0.40 in/1M$4.00 out/1M60 tok/s131Kfp816m agoNovita AIzero-retention$0.98 in/1M$3.95 out/1M41 tok/s131Kbf1616m ago
Full specs, hardware verdicts and benchmarks on theQwen3 VL 235B A22B Thinking model page →