Compatibility check
Can I run Qwen2.5 7B Instruct on GeForce GTX 1660 SUPER?
Partially — with CPU offload.best quant:
4.8 GB of weights, plus 1.5 GB for the software that runs it and the smallest conversation it can hold, comes to 6.3 GB against the 5.5 GB this 6 GB device leaves free.
Decode speed
—
Usable context
—
Memory
6 GB
Spills to system RAM
Spills to system RAM
Too large
Plan B — rent it
Qwen2.5 7B Instruct is hosted by 3 providers from $0.20 per 1M output tokens (Phala).Compare providers →