Hardware Benchmarks

Best Local LLM Performance for RX 7900 Xt 7900 XTX 7900 Gre 7900m

Compare RX 7900 Xt 7900 XTX 7900 Gre 7900m benchmark results and see which models actually earn the best completed quality scores.

ShareRedditX

Hardware Overview

GPU-specific specifications for local LLM planning.

AMD Radeon RX 7900 XT/7900 XTX/7900 GRE/7900M
VRAM
23GB

Recommended Models

Models we recommend for RX 7900 Xt 7900 XTX 7900 Gre 7900m.

Fastest & Best quality & Largest context

Qwen3.8-27B-Q4_K_XL

27B

75.6 tok/sQuality 86.5Context 65,536 tokens

Highest token generation speed at 75.6 tok/s. Highest completed quality score at 86.5. Largest context that fits in 23GB VRAM, at 65,536 tokens.

Frequently asked questions

Which models actually run best on the RX 7900 Xt 7900 XTX 7900 Gre 7900m, by task, from community benchmark data.

What is the best local model for coding on the RX 7900 Xt 7900 XTX 7900 Gre 7900m?
On the RX 7900 Xt 7900 XTX 7900 Gre 7900m, Qwen3.8-27B-Q4_K_XL ranks first for coding in our benchmarks (81.5/100 at ~71.0 tok/s).
What is the best local model for agentic workflows on the RX 7900 Xt 7900 XTX 7900 Gre 7900m?
On the RX 7900 Xt 7900 XTX 7900 Gre 7900m, Qwen3.8-27B-IQ3_XXS ranks first for agentic workflows in our benchmarks (87.0/100 at ~58.0 tok/s). Qwen3.8-27B-Q4_K_XL is faster (~71.0 tok/s) and still scores well (83.4/100) β€” a good pick if you'd rather trade a little quality for speed.

Benchmark Results

Completed quality-ranked benchmark results for AMD Radeon RX 7900 XT/7900 XTX/7900 GRE/7900M.