Models /Qwen3 235B A22B Thinking 2507 /Where to run
Provider guide

Where to run Qwen3 235B A22B Thinking 2507

8 live offers tracked — output prices vary 2.0× between the cheapest and most expensive host, so the provider choice matters as much as the model choice.

CHEAPEST
$0.23 / $2.30
per 1M tokens in / out
FASTEST MEASURED
69 tok/s
measured throughput · $0.23 input / $2.30 output per million tokens
Alibaba Cloud$0.15 in/1M$1.50 out/1M47 tok/s131K6d agoDeepInfra$0.23 in/1M$2.30 out/1M262Kfp86h agoOpenRouterCHEAPEST$0.23 in/1M$2.30 out/1M262K17m agoAlibaba Cloud$0.23 in/1M$2.30 out/1M69 tok/s131Kfp816m agoNovita AI$0.30 in/1M$3.00 out/1M131K6h agoDeepInfrazero-retention$0.23 in/1M$2.30 out/1M60 tok/s262Kfp816m agoNovita AIzero-retention$0.30 in/1M$3.00 out/1M33 tok/s131Kfp816m agoVenice AIzero-retention$0.45 in/1M$3.50 out/1M62 tok/s128Kfp816m ago
Full specs, hardware verdicts and benchmarks on theQwen3 235B A22B Thinking 2507 model page →