Models /Qwen2.5 7B Instruct /Where to run
Provider guide

Where to run Qwen2.5 7B Instruct

3 live listings tracked — output prices vary 2.9× between the cheapest and most expensive host, so the provider choice matters as much as the model choice.

CHEAPEST
$0.10 / $0.20
per 1M tokens in / out
FASTEST MEASURED
28 tok/s
measured throughput · $0.10 input / $0.20 output per million tokens
Novita AI$0.070 in/1M$0.070 out/1M—32K—21 days agoOpenRouter$0.10 in/1M$0.20 out/1M—33K—4 hours agoPhalazero-retentionCHEAPEST$0.10 in/1M$0.20 out/1M28 tok/s33K—4 hours ago
Full specs, hardware verdicts and benchmarks on the Qwen2.5 7B Instruct model page →