Hardware Benchmarks

Best Local LLM Performance for 2× RX 9050 9060 Xt

Compare 2× RX 9050 9060 Xt benchmark results and see which models actually earn the best completed quality scores.

ShareRedditX

Hardware Overview

GPU-specific specifications for local LLM planning.

2× AMD Radeon RX 9050 / 9060 XT
VRAM
30GB
Architecture
Multi-GPU

Recommended Models

Models we recommend for 2× RX 9050 9060 Xt.

Fastest & Largest context

gemma4:12b-it-qat

11.9B

33.5 tok/sContext 65.536 tokens

Highest token generation speed at 33.5 tok/s. Largest context that fits in 30GB VRAM, at 65.536 tokens.

Best quality

qwen3.8-flash-next:125b-a6b-q4_K_M

176.9B

13.4 tok/sQuality 80.5

Highest completed quality score at 80.5.

Frequently asked questions

Which models actually run best on the 2× RX 9050 9060 Xt, by task, from community benchmark data.

What is the best local model for coding on the 2× RX 9050 9060 Xt?
On the 2× RX 9050 9060 Xt, qwen3.8-flash-next:125b-a6b-q4_K_M ranks first for coding in our benchmarks (77.8/100 at ~13.4 tok/s).
What is the best local model for agentic workflows on the 2× RX 9050 9060 Xt?
On the 2× RX 9050 9060 Xt, gemma4:12b-it-qat ranks first for agentic workflows in our benchmarks (82.5/100 at ~33.5 tok/s).

Benchmark Results

Completed quality-ranked benchmark results for 2× AMD Radeon RX 9050 / 9060 XT.