Fastest & Largest context
qwen3:8b
8.2B
82.4 tok/sContext 8.192 tokens
Highest token generation speed at 82.4 tok/s. Largest context that fits in 15GB VRAM, at 8.192 tokens.
Hardware Benchmarks
Compare RX 9070 Xt benchmark results and see which models actually earn the best completed quality scores.
GPU-specific specifications for local LLM planning.
Models we recommend for RX 9070 Xt.
Fastest & Largest context
8.2B
Highest token generation speed at 82.4 tok/s. Largest context that fits in 15GB VRAM, at 8.192 tokens.
Best quality
27.3B
Highest completed quality score at 85.0.
Which models actually run best on the RX 9070 Xt, by task, from community benchmark data.
Completed quality-ranked benchmark results for AMD Radeon RX 9070 XT.