Compatibility check
Can I run Qwen3 30B A3B Thinking 2507 on Radeon RX 7900 XT?
Partially — with CPU offload.best quant:
19.2 GB of weights, plus 2 GB for the software that runs it and the smallest conversation it can hold, comes to 21.2 GB against the 18.8 GB this 20 GB device leaves free.
Decode speed
—
Usable context
—
Memory
20 GB
Spills to system RAM
Spills to system RAM
Too large
Plan B — rent it
Qwen3 30B A3B Thinking 2507 is hosted by 2 providers from $2.40 per 1M output tokens (Alibaba Cloud).Compare providers →