Hardware Benchmarks

Best Local LLM Performance for RX 7900 XTX

Compare RX 7900 XTX benchmark results and see which models actually earn the best completed quality scores.

ShareRedditX

Hardware Overview

GPU-specific specifications for local LLM planning.

AMD RX 7900 XTX
VRAM
24GB
Architecture
RDNA 3
Memory Bandwidth
960.0 GB/s

Recommended Models

Models we recommend for RX 7900 XTX.

Fastest

hf.co/LiquidAI/LFM2.5-2.6B-GGUF:Q8_0

2.7B

169.3 tok/s

Highest token generation speed at 169.3 tok/s.

Best quality

qwen3.8:27b-mtp-q4_K_M

27.3B

41.7 tok/sQuality 87.5

Highest completed quality score at 87.5.

Largest context

openai/gpt-oss-20b

20B

108.3 tok/sContext 256.000 tokens

Largest context that fits in 24GB VRAM, at 256.000 tokens.

Frequently asked questions

Which models actually run best on the RX 7900 XTX, by task, from community benchmark data.

What is the best local model for coding on the RX 7900 XTX?
On the RX 7900 XTX, qwen3.8:27b-q4_K_M ranks first for coding in our benchmarks (80.7/100 at ~34.5 tok/s).
What is the best local model for agentic workflows on the RX 7900 XTX?
On the RX 7900 XTX, muse-glimmer:latest ranks first for agentic workflows in our benchmarks (93.0/100 at ~34.5 tok/s). hf.co/ornith-ai/Ornith-1.0-9B-GGUF:Q8_0 is faster (~68.7 tok/s) and still scores well (91.5/100) — a good pick if you'd rather trade a little quality for speed.

Benchmark Results

Completed quality-ranked benchmark results for AMD RX 7900 XTX.