Models /Qwen3 VL 30B A3B Thinking /Can I run it?
Compatibility check

Can I run Qwen3 VL 30B A3B Thinking on Radeon RX 7900 XT?

Partially — with CPU offload.best quant:

19.6 GB of weights, plus 2 GB for the software that runs it and the smallest conversation it can hold, comes to 21.6 GB against the 18.8 GB this 20 GB device leaves free.

Decode speed
—
Usable context
—
Memory
20 GB
Spills to system RAM
Spills to system RAM
Too large
What is quantisation? →
Plan B — rent it

Qwen3 VL 30B A3B Thinking is hosted by 4 providers from $2.40 per 1M output tokens (OpenRouter).Compare providers →

Qwen3 VL 30B A3B Thinking
31.1B / 3B · full specs & benchmarks →
Radeon RX 7900 XT
20 GB · everything it runs →