Fastest & Largest context
/home/drzewkam/.lmstudio/models/Qwen3.8-27B-ROCMFP2.gguf
27B
5167.3 tok/sContext 65,536 tokens
Highest token generation speed at 5167.3 tok/s. Largest context that fits in 11GB VRAM, at 65,536 tokens.
Hardware Benchmarks
Compare RX 6700 6700 Xt 6750 Xt 6800m 6850m Xt benchmark results and see which models actually earn the best completed quality scores.
GPU-specific specifications for local LLM planning.
Models we recommend for RX 6700 6700 Xt 6750 Xt 6800m 6850m Xt.
Fastest & Largest context
27B
Highest token generation speed at 5167.3 tok/s. Largest context that fits in 11GB VRAM, at 65,536 tokens.
Best quality
9B
Highest completed quality score at 76.7.
Completed quality-ranked benchmark results for AMD Radeon RX 6700/6700 XT/6750 XT / 6800M/6850M XT.