Which AI models run on a AMD Radeon RX 7900 XT?

The AMD Radeon RX 7900 XT has 20 GB of VRAM. Here are the popular AI models it can run locally (4,096-token context, ~32.0 GB system RAM assumed), ranked by popularity.

See also: Best GPU for running local LLMs.

VRAM
20 GB
Vendor
AMD
Fits in VRAM
35 models
Assumed RAM
32.0 GB

The AMD Radeon RX 7900 XT comes with 20 GB of VRAM. Among the popular GGUF models we track, it can run 35 of them entirely in VRAM — including Qwen3-Coder-30B-A3B-Instruct-GGUF, gpt-oss-20b-GGUF, Qwen-AgentWorld-35B-A3B-GGUF.

With 20 GB you can typically run a 14B at high quality, or a 24–27B at 4-bit. Which quantization is best depends on the exact model and your context length. For a full shortlist, see the best LLM for 16 GB of VRAM.

Larger models such as Laguna-S-2.1-GGUF still run on a AMD Radeon RX 7900 XT but require offloading part of the model to system RAM, which lowers speed. Models that exceed both VRAM and RAM are not listed.

New to this? Read: How much VRAM do you need?

35 fit fully in VRAM · 3 run with offload

ModelSize Quant.Quality MemorySpeed~ Verdict
unsloth/Qwen3-Coder-30B-A3B-Instruct-GGUF 30.53B Q4_1 Very good
19.05 GB
22.4 t/s Fits in VRAM
unsloth/gpt-oss-20b-GGUF 20.91B F16 Very good
13.74 GB
31.1 t/s Fits in VRAM
unsloth/Qwen-AgentWorld-35B-A3B-GGUF 34.66B IQ4_NL Fair
17.75 GB
23.7 t/s Fits in VRAM
MaziyarPanahi/Qwen3-4B-GGUF 4.02B GGUF Excellent
8.86 GB
53.3 t/s Fits in VRAM
bartowski/Meta-Llama-3.1-8B-Instruct-GGUF 8.03B Q8_0 Excellent
9.23 GB
50.3 t/s Fits in VRAM
janhq/Jan-v3.5-4B-gguf 4.41B GGUF Excellent
9.59 GB
48.6 t/s Fits in VRAM
Qwen/Qwen3-8B-GGUF 8.19B Q8_0 Excellent
9.47 GB
49.3 t/s Fits in VRAM
MaziyarPanahi/Qwen3-0.6B-GGUF 0.75B GGUF Excellent
2.64 GB
284.6 t/s Fits in VRAM
MaziyarPanahi/Qwen3-14B-GGUF 14.77B Q6_K Excellent
12.71 GB
35.4 t/s Fits in VRAM
MaziyarPanahi/Qwen3-1.7B-GGUF 2.03B GGUF Excellent
5.03 GB
105.5 t/s Fits in VRAM
MaziyarPanahi/Qwen3-32B-GGUF 32.76B Q3_K_L Good
17.94 GB
24.8 t/s Fits in VRAM
MaziyarPanahi/Qwen3-30B-A3B-GGUF 30.53B Q4_K_M Good
18.46 GB
23.1 t/s Fits in VRAM
bartowski/Qwen2.5-7B-Instruct-GGUF 7.62B F16 Excellent
15.21 GB
28.2 t/s Fits in VRAM
unsloth/Ornith-1.0-35B-GGUF IQ4_NL Excellent
17.75 GB
23.7 t/s Fits in VRAM
Qwen/Qwen2.5-3B-Instruct-GGUF 3.09B GGUF Excellent
7.27 GB
63.2 t/s Fits in VRAM
LiquidAI/LFM2.5-2.6B-GGUF 2.7B BF16 Excellent
5.89 GB
79.5 t/s Fits in VRAM
LiquidAI/LFM2.5-1.2B-Instruct-GGUF 1.17B BF16 Excellent
3.03 GB
183.3 t/s Fits in VRAM
bartowski/Kwaipilot_KAT-Coder-V2.5-Dev-GGUF 34.66B Q4_0 Good
19.45 GB
21.5 t/s Fits in VRAM
Qwen/Qwen2.5-1.5B-Instruct-GGUF 1.54B GGUF Excellent
4.23 GB
120.6 t/s Fits in VRAM
unsloth/Qwen3-Coder-Next-GGUF 79.67B TQ1_0 Very low
18.82 GB
22.7 t/s Fits in VRAM
Qwen/Qwen2.5-Coder-7B-Instruct-GGUF 7.62B GGUF Excellent
15.21 GB
28.2 t/s Fits in VRAM
MaziyarPanahi/Qwen3-4B-Instruct-2507-GGUF 4.02B GGUF Excellent
8.86 GB
53.3 t/s Fits in VRAM
MaziyarPanahi/Meta-Llama-3-8B-Instruct-GGUF 8.03B GGUF Excellent
16.24 GB
26.7 t/s Fits in VRAM
MaziyarPanahi/Mistral-7B-Instruct-v0.3-GGUF 7.25B GGUF Excellent
14.8 GB
29.6 t/s Fits in VRAM
Qwen/Qwen2.5-0.5B-Instruct-GGUF 0.49B GGUF Excellent
2.03 GB
339.1 t/s Fits in VRAM
MaziyarPanahi/gemma-3-4b-it-GGUF 4.3B GGUF Excellent
8.38 GB
55.3 t/s Fits in VRAM
MaziyarPanahi/Phi-3.5-mini-instruct-GGUF 3.82B Q8_0 Excellent
6.08 GB
105.8 t/s Fits in VRAM
lmstudio-community/Llama-3.2-3B-Instruct-GGUF 3.21B Q8_0 Excellent
4.29 GB
125.5 t/s Fits in VRAM
MaziyarPanahi/Llama-3-8B-Instruct-32k-v0.1-GGUF 8.03B GGUF Excellent
16.24 GB
26.7 t/s Fits in VRAM
MaziyarPanahi/Mistral-Small-24B-Instruct-2501-GGUF 23.57B Q6_K Excellent
19.44 GB
22.2 t/s Fits in VRAM
MaziyarPanahi/Yi-1.5-6B-Chat-GGUF 6.06B Q6_K Excellent
5.68 GB
86.3 t/s Fits in VRAM
MaziyarPanahi/Llama-3.2-1B-Instruct-GGUF 1.24B GGUF Excellent
3.3 GB
173.2 t/s Fits in VRAM
MaziyarPanahi/Mistral-Nemo-Instruct-2407-GGUF 12.25B Q8_0 Excellent
13.55 GB
33.0 t/s Fits in VRAM
unsloth/GLM-4.7-Flash-GGUF 31.22B Q4_1 Good
18.68 GB
22.6 t/s Fits in VRAM
MaziyarPanahi/WizardLM-2-7B-GGUF 7.24B GGUF Excellent
14.74 GB
29.7 t/s Fits in VRAM
unsloth/Laguna-S-2.1-GGUF 117.56B Q3_K_M Fair
51.3 GB
1.0 t/s Offload
MaziyarPanahi/Mixtral-8x22B-v0.1-GGUF 140.62B Q2_K Low
50.2 GB
1.0 t/s Offload
MaziyarPanahi/Llama-3.3-70B-Instruct-GGUF 70.55B Q5_K_M Very good
48.74 GB
1.1 t/s Offload

"Fits in VRAM" = fast, fully on GPU. "Offload" = part on system RAM, slower. Speed is a rough estimate.

Frequently asked questions

How much VRAM does the AMD Radeon RX 7900 XT have?

The AMD Radeon RX 7900 XT has 20 GB of VRAM, which determines how large a model it can run entirely on the GPU.

What is the best LLM to run on a AMD Radeon RX 7900 XT?

Among popular models, unsloth/Qwen3-Coder-30B-A3B-Instruct-GGUF runs well on a AMD Radeon RX 7900 XT using the Q4_1 quantization (about 19.05 GB). With 20 GB you can generally run a 14B at high quality, or a 24–27B at 4-bit. Larger models trade speed for capability via RAM offloading. See the best LLM for 16 GB of VRAM.

Can a AMD Radeon RX 7900 XT run a 7–8B model?

Yes. A 7–8B model like Meta-Llama-3.1-8B-Instruct-GGUF fits entirely in the 20 GB of a AMD Radeon RX 7900 XT (Q8_0).

Can a AMD Radeon RX 7900 XT run a 13–14B model?

Yes. A 13–14B model like Qwen3-14B-GGUF fits entirely in the 20 GB of a AMD Radeon RX 7900 XT (Q6_K).

Can a AMD Radeon RX 7900 XT run a 70B model?

Yes. A 70B model like Qwen3-Coder-Next-GGUF fits entirely in the 20 GB of a AMD Radeon RX 7900 XT (TQ1_0).

Another graphics card

NVIDIA RTX 5090 32 GBNVIDIA RTX 4090 24 GBNVIDIA RTX 3090 Ti 24 GBNVIDIA RTX 3090 24 GBNVIDIA RTX 5080 16 GBNVIDIA RTX 5070 Ti 16 GBNVIDIA RTX 4080 Super 16 GBNVIDIA RTX 4080 16 GBNVIDIA RTX 4070 Ti Super 16 GBNVIDIA RTX 5060 Ti 16 GB 16 GBNVIDIA RTX 4060 Ti 16 GB 16 GBNVIDIA RTX 5070 12 GBNVIDIA RTX 4070 Ti 12 GBNVIDIA RTX 4070 Super 12 GBNVIDIA RTX 4070 12 GBNVIDIA RTX 3080 Ti 12 GBNVIDIA RTX 3060 12 GB 12 GBNVIDIA RTX 2080 Ti 11 GBNVIDIA RTX 3080 10 GBNVIDIA RTX 5060 8 GBNVIDIA RTX 4060 Ti 8 GB 8 GBNVIDIA RTX 4060 8 GBNVIDIA RTX 3070 Ti 8 GBNVIDIA RTX 3070 8 GBNVIDIA RTX 3060 Ti 8 GBNVIDIA RTX 2080 Super 8 GBNVIDIA RTX 2070 Super 8 GBNVIDIA RTX 2060 Super 8 GBNVIDIA RTX 3050 8 GBNVIDIA RTX 2060 6 GBNVIDIA GTX 1660 Ti 6 GBNVIDIA GTX 1660 Super 6 GBNVIDIA GTX 1660 6 GBNVIDIA GTX 1650 4 GBNVIDIA RTX 5090 Laptop 24 GBNVIDIA RTX 5080 Laptop 16 GBNVIDIA RTX 5070 Ti Laptop 12 GBNVIDIA RTX 5070 Laptop 8 GBNVIDIA RTX 5060 Laptop 8 GBNVIDIA RTX 5050 Laptop 8 GBNVIDIA RTX 4090 Laptop 16 GBNVIDIA RTX 4080 Laptop 12 GBNVIDIA RTX 4070 Laptop 8 GBNVIDIA RTX 4060 Laptop 8 GBNVIDIA RTX 4050 Laptop 6 GBNVIDIA RTX 3080 Ti Laptop 16 GBNVIDIA RTX 3070 Ti Laptop 8 GBNVIDIA RTX 3070 Laptop 8 GBNVIDIA RTX 3060 Laptop 6 GBNVIDIA RTX 3050 Ti Laptop 4 GBNVIDIA RTX 3050 Laptop 4 GBAMD Radeon RX 7900 XTX 24 GBAMD Radeon RX 7800 XT 16 GBAMD Radeon RX 7600 XT 16 GBAMD Radeon RX 6950 XT 16 GBAMD Radeon RX 6800 XT 16 GBAMD Radeon RX 6800 16 GBAMD Radeon RX 7700 XT 12 GBAMD Radeon RX 6750 XT 12 GBAMD Radeon RX 6700 XT 12 GBAMD Radeon RX 7600 8 GBAMD Radeon RX 6650 XT 8 GBAMD Radeon RX 6600 8 GBIntel Arc A770 16 GBIntel Arc B580 12 GBIntel Arc A750 8 GB