Hardware Benchmarks

Best Local LLM Performance for RX 9070 Xt

Compare RX 9070 Xt benchmark results and see which models actually earn the best completed quality scores.

ShareRedditX

Hardware Overview

GPU-specific specifications for local LLM planning.

AMD Radeon RX 9070 XT
VRAM
15GB

Recommended Models

Models we recommend for RX 9070 Xt.

Fastest & Largest context

qwen3:8b

8.2B

82.4 tok/sContext 8.192 tokens

Highest token generation speed at 82.4 tok/s. Largest context that fits in 15GB VRAM, at 8.192 tokens.

Best quality

qwen3.8:27b

27.3B

15.0 tok/sQuality 85.0

Highest completed quality score at 85.0.

Frequently asked questions

Which models actually run best on the RX 9070 Xt, by task, from community benchmark data.

What is the best local model for coding on the RX 9070 Xt?
On the RX 9070 Xt, qwen3.8:27b ranks first for coding in our benchmarks (80.2/100 at ~15.0 tok/s).
What is the best local model for agentic workflows on the RX 9070 Xt?
On the RX 9070 Xt, qwen3.8:27b ranks first for agentic workflows in our benchmarks (81.1/100 at ~15.0 tok/s).

Benchmark Results

Completed quality-ranked benchmark results for AMD Radeon RX 9070 XT.