Models /Qwen3 VL 8B Instruct /Where to run
Provider guide

Where to run Qwen3 VL 8B Instruct

4 live listings tracked — output prices vary 1.6× between the cheapest and most expensive host, so the provider choice matters as much as the model choice.

CHEAPEST
$0.12 / $0.46
per 1M tokens in / out
FASTEST MEASURED
44 tok/s
measured throughput · $0.12 input / $0.46 output per million tokens
OpenRouterCHEAPEST$0.12 in/1M$0.46 out/1M—262K—4 hours agoAlibaba Cloud$0.12 in/1M$0.46 out/1M44 tok/s131K—4 hours agoNovita AI$0.080 in/1M$0.50 out/1M—131K—21 days agoParasailzero-retention$0.25 in/1M$0.75 out/1M10 tok/s262Kbf1610 hours ago
Full specs, hardware verdicts and benchmarks on the Qwen3 VL 8B Instruct model page →