Hardware Benchmarks

Best Local LLM Performance for RX 7900 Gre

Compare RX 7900 Gre benchmark results and see which models actually earn the best completed quality scores.

ShareRedditX

Hardware Overview

GPU-specific specifications for local LLM planning.

AMD Radeon RX 7900 GRE
VRAM
16GB

Recommended Models

Models we recommend for RX 7900 Gre.

Fastest & Best quality & Largest context

qwen3.8-27b-gsq-rco@iq2_xs

27B

33.9 tok/sQuality 87.3Context 32,768 tokens

Highest token generation speed at 33.9 tok/s. Highest completed quality score at 87.3. Largest context that fits in 16GB VRAM, at 32,768 tokens.

Frequently asked questions

Which models actually run best on the RX 7900 Gre, by task, from community benchmark data.

What is the best local model for coding on the RX 7900 Gre?
On the RX 7900 Gre, qwen3.8-27b-gsq-rco@iq2_xs ranks first for coding in our benchmarks (78.0/100 at ~31.2 tok/s).
What is the best local model for agentic workflows on the RX 7900 Gre?
On the RX 7900 Gre, qwen3.8-27b-gsq-rco@iq2_xs ranks first for agentic workflows in our benchmarks (64.6/100 at ~31.2 tok/s).

Benchmark Results

Completed quality-ranked benchmark results for AMD Radeon RX 7900 GRE.