Models /Qwen3 VL 30B A3B Thinking /Where to run
Provider guide

Where to run Qwen3 VL 30B A3B Thinking

4 live listings tracked — output prices vary 2.4× between the cheapest and most expensive host, so the provider choice matters as much as the model choice.

CHEAPEST
$0.20 / $2.40
per 1M tokens in / out
FASTEST MEASURED
113 tok/s
measured throughput · $0.20 input / $2.40 output per million tokens
Novita AI$0.20 in/1M$1.00 out/1M—131K—21 days agoSiliconFlowzero-retention$0.29 in/1M$1.00 out/1M83 tok/s262Kfp83 hours agoOpenRouterCHEAPEST$0.20 in/1M$2.40 out/1M—262K—3 hours agoAlibaba Cloud$0.20 in/1M$2.40 out/1M113 tok/s131K—3 hours ago
Full specs, hardware verdicts and benchmarks on the Qwen3 VL 30B A3B Thinking model page →