AMD Radeon PRO W7900

The AMD Radeon PRO W7900 has 48 GB VRAM and 864 GB/s memory bandwidth. It can run 57 of our 94 tracked models natively in VRAM at 8k context.

With 48 GB GDDR6, the AMD Radeon PRO W7900 is a workstation-tier GPU that can run 57 models natively. It handles 70B-class models at Q4 quantization.

The AMD Radeon PRO W7900 is AMD's flagship RDNA 3 workstation GPU, delivering 48GB ECC-capable GDDR6 on a 384-bit bus at 864 GB/s with 96MB of Infinity Cache and 6,144 stream processors. It can hold 34B models at Q8_0 and 70B models at Q4_K_M in VRAM, matching the NVIDIA RTX 6000 Ada on capacity at comparable bandwidth, making it AMD's most capable single-GPU workstation inference platform. ROCm applies on Linux; use the Vulkan backend on Windows.

AMD Radeon PRO W7900: 2023 RDNA 3 flagship workstation GPU with 48GB ECC-capable GDDR6 on a 384-bit bus at 864 GB/s, 96MB Infinity Cache, 6,144 stream processors.

70B models fit at Q4_K_M; 34B models fit at Q8_0 entirely in VRAM, matching the NVIDIA RTX 6000 Ada on capacity at comparable bandwidth. ~25-35 t/s for 7B Q4.

ROCm on Linux for full acceleration; Vulkan backend on Windows. AMD's most capable single-GPU workstation option for local LLM inference.

VendorAMD
ArchitectureRDNA 3
VRAM48 GB
Memory typeGDDR6
Memory bandwidth864 GB/s
Compute backendROCM
TierWorkstation
Released2023
Models (native)57 / 94
Models (offload)8 / 94
Software: ROCm is Linux-only; on Windows use the Vulkan backend instead. Requires llama.cpp compiled with ROCm support.

Popular models for this GPU

Models this GPU runs natively in VRAM (57)

Show 52 more

Models that fit with CPU offload (8)

These use system RAM for layers that don't fit in VRAM, so expect much slower inference.

Too large for this GPU (29)

Frequently asked questions

How much VRAM does the AMD Radeon PRO W7900 have?
The AMD Radeon PRO W7900 has 48 GB of GDDR6 with 864 GB/s memory bandwidth.
What is the AMD Radeon PRO W7900 best for?
With 48 GB of VRAM, the AMD Radeon PRO W7900 is ideal for running 70B-class models at Q4 quantization and large MoE models, a workstation sweet spot for local inference.
What LLMs can the AMD Radeon PRO W7900 run locally?
The AMD Radeon PRO W7900 can run 57 of the 94 open-weight models tracked by CanItRun natively in VRAM at 8k context. Top options include: Qwen 3.8 27B at Q8_0, Muse Glimmer 30B at Q8_0, Ornith 1.5 9B at FP32.
Can the AMD Radeon PRO W7900 run Gemma 4 31B?
Yes. The AMD Radeon PRO W7900 runs Gemma 4 31B natively in VRAM at Q8_0 quantization, achieving approximately 16.4 tokens per second.
Can the AMD Radeon PRO W7900 run Qwen 3.6 27B?
Yes. The AMD Radeon PRO W7900 runs Qwen 3.6 27B natively in VRAM at Q8_0 quantization, achieving approximately 19.2 tokens per second.
Can the AMD Radeon PRO W7900 run Qwen3 8B?
Yes. The AMD Radeon PRO W7900 runs Qwen3 8B natively in VRAM at FP32 quantization, achieving approximately 16.9 tokens per second.