self-hosted/ai
§01·compatibility · /check

Llama 3.1 8B on RX 7900 XTX

Yes — Llama 3.1 8B runs on the RX 7900 XTX (24 GB). Fastest community-measured result: 51.3 tokens/s.

✓ runsllmactiveamd series24GB VRAM
step-by-step recipe for this pair

Llama 3.1 8B on Radeon RX 7900 XTX: Local Chat via Ollama or llama.cpp HIP (ROCm) GGUF

llmbeginner8GB+
model
name
Llama 3.1 8B
slug
llama-3-1-8b
vertical
llm
status
active
open detail ↗
gpu
name
RX 7900 XTX
slug
rx-7900-xtx
vram
24 GB
series
amd
open detail ↗
§02·benchmarks
TaskQuantSpeedVRAMWorksConfidenceSourceVerified
llmQ4_K - Medium51.3tokens/s✓localscore.ai· web2026-05-15
§03·how it was measured
  • llmQ4_K - Medium51.3 tokens/s

    Prompt Speed: 870 tokens/s; TTFT: 1.44sec

    recorded from localscore.ai ↗ · verified 2026-05-15

§04·more Llama 3.1 8B recipes
§05·common questions
Can you run Llama 3.1 8B on RX 7900 XTX?

Yes — Llama 3.1 8B runs on the RX 7900 XTX (24 GB). Fastest community-measured result: 51.3 tokens/s.

Which quantizations have been tested for Llama 3.1 8B on RX 7900 XTX?

Q4_K - Medium — measured in community benchmarks.

How fast is Llama 3.1 8B on RX 7900 XTX?

Up to 51.3 tokens/s (llm), the fastest community-measured result.

Are there step-by-step instructions for Llama 3.1 8B on RX 7900 XTX?

Yes — a step-by-step recipe documents Llama 3.1 8B on the RX 7900 XTX, linked at the top of this page.