作者: Jakub Rusinowski · 最后更新: 2026年7月12日
Entry-level RDNA 4 card at $299. 8 GB VRAM covers 7–8B models in Q4. Same Navi 44 chip and bandwidth as the 16GB variant.
| VRAM | 8 GB |
| Memory Bandwidth | 320 GB/s |
| TDP | 150 W |
| Architecture | RDNA 4 Navi 44 |
| Release Year | 2025 |
| MSRP at Launch | $299 |
| Inference Speed (Llama 3.1 8B Q4_K_M) | 24–49 tok/s (estimated) |
| Inference Speed (Llama 3.3 70B Q4_K_M) | Does not fit — needs ~44 GB of 8 GB usable |
or compare on Vast.ai from $0.35/hr (typical low · varies)
As an Amazon Associate we earn from qualifying purchases. Cloud GPU links are referral links — we may earn a commission at no extra cost to you.
All models below run comfortably in 8 GB VRAM with Q4_K_M quantization.
| Llama 3.1 Family | Llama 3.1 8B Instruct · 6 GB VRAM · Q4_K_M · ollama run llama3.1 |
| Llama 3.2 Family | Llama 3.2 11B Vision Instruct · 7 GB VRAM · Q4_K_M · llama-3-2 |
| Qwen 2.5 Family | Qwen 2.5 7B Instruct · 5 GB VRAM · Q4_K_M · ollama run qwen2.5:7b |
| Gemma 3 | Gemma 3 4B Instruct · 3 GB VRAM · Q4_K_M · ollama run gemma3:4b |
| Phi-4 Mini | Phi-4 Mini (3.8B) · 3 GB VRAM · Q4_K_M · ollama run phi4-mini |
| SmolLM2 | SmolLM2 1.7B Instruct · 2 GB VRAM · Q4_K_M · ollama run smollm2:1.7b |
| Ministral 3 | Ministral 3 8B · 6 GB VRAM · Q4_K_M · ollama run ministral-3:8b |
| IBM Granite 4.2 | Granite 4.2 8B · 6 GB VRAM · Q4_K_M · ollama run granite4.2:8b |
Install Ollama then run the recommended model for this GPU:
ollama run llama3.1:8b
Yes — the AMD Radeon RX 9060 XT 8GB has 8 GB VRAM and runs Entry-level RDNA 4 card at $299. 8 GB VRAM covers 7–8B models in Q4. Same Navi 44 chip and bandwidth as the 16GB variant
The AMD Radeon RX 9060 XT 8GB is estimated to run Llama 3.1 8B at 24–49 tok/s with Q4_K_M quantization. Llama 3.3 70B does not fit: it needs about 44 GB against 8 GB usable. These are modelled estimates, not measurements — see /en/methodology.
With 8 GB you can run: Llama 3.1 Family, Llama 3.2 Family, Qwen 2.5 Family, Gemma 3, Phi-4 Mini. Use Ollama for the easiest setup: ollama run llama3.1:8b.
← All GPU Reviews | All Hardware | Check Your Hardware | Full Benchmarks | Can I Run It?