AMDDesktop GPUPrevious generation

AMD Radeon RX 6900 XT for local LLMs

Written by Jakub Rusinowski · Last updated

With 16 GB of GDDR6 at 512 GB/s, the RX 6900 XT runs 74 catalogued models at Q4_K_M with 8K context. The largest that fits is OLMo 2 13B Instruct (~15.8 GB), and the top pick is EuroLLM 22B.

16 GB of GDDR6, 512 GB/s: the same memory system as the RX 6800 XT, so the same decode speed on a model that fits. No speed calibration for RDNA 2: fit verdicts only.

Models that run on the RX 6900 XT

Q4_K_M, 8K context, 16 GB usable. Ranked by quality and speed.

ModelVRAMSpeed
EuroLLM 22B
EuroLLM
14.1 GB—
GPT-OSS 20B
GPT-OSS
13.8 GB—
InternLM 3 20B Instruct
InternLM 3
12.9 GB—
Cosmos 3 Nano
Cosmos 3
10.5 GB—
StarCoder 2 15B
StarCoder 2
10.8 GB—
Qwen 3 14B
Qwen 3
11.1 GB—
Phi-4 (14B)
Phi-4 Family
10.9 GB—
Cogito v1 14B
Cogito v1
10.9 GB—
Ministral 3 14B
Ministral 3
9.3 GB—
DeepSeek R1 Distill Qwen 14B
DeepSeek R1
10.6 GB—
Gemma 4 12B (Unified)
Gemma 4
8 GB—
Bielik PL 11B v3.0 Instruct
Bielik
7.4 GB—
Showing 12 of 74

Buy it or rent the same memory

Buy the card, or rent a GPU with the same memory by the hour to try models first.

Speed vs other GPUs

Llama 3.1 8B, Q4_K_M. Estimated ranges. How this is calculated

Specifications

Specs last updated 2026-10-07.

Memory
16 GB GDDR6
Memory bandwidth
512 GB/s
Memory bus
256-bit
Architecture
RDNA 2 Navi 21
Series
RX 6000-series
Board power
300 W
Release year
2020
Compute backends
ROCM, VULKAN
Usable for models
16 GB
Best forused 16 GBVulkan

Similar GPUs

Frequently asked questions

Can the AMD Radeon RX 6900 XT run local LLMs?

Yes. With 16 GB (16 GB usable by a model) it runs 74 of the catalogued models at Q4_K_M with 8K context; the largest is OLMo 2 13B Instruct, needing about 15.8 GB.

How fast is the AMD Radeon RX 6900 XT for AI inference?

This site has no speed calibration for the RDNA 2 Navi 21 architecture, so it publishes fit verdicts for this card but no tokens-per-second estimate. Llama 3.3 70B does not fit: it needs about 44 GB against 16 GB usable.

What LLMs can I run on 16 GB?

Among the best that fit: EuroLLM 22B, GPT-OSS 20B, InternLM 3 20B Instruct, Cosmos 3 Nano, StarCoder 2 15B.