Fastest & Largest context
gemma4:12b-it-qat
11.9B
33.5 tok/sContext 65.536 tokens
Highest token generation speed at 33.5 tok/s. Largest context that fits in 30GB VRAM, at 65.536 tokens.
Hardware Benchmarks
Compare 2× RX 9050 9060 Xt benchmark results and see which models actually earn the best completed quality scores.
GPU-specific specifications for local LLM planning.
Models we recommend for 2× RX 9050 9060 Xt.
Fastest & Largest context
11.9B
Highest token generation speed at 33.5 tok/s. Largest context that fits in 30GB VRAM, at 65.536 tokens.
Best quality
176.9B
Highest completed quality score at 80.5.
Which models actually run best on the 2× RX 9050 9060 Xt, by task, from community benchmark data.
Completed quality-ranked benchmark results for 2× AMD Radeon RX 9050 / 9060 XT.