Fastest & Largest context
DavidAU/LFM2.5-2.6B-Qwen3.8-Turbo-Brilliance-Power-X12-NEO-MAX-GGUF
2.6B
176.0 tok/sContext 8,192 tokens
Highest token generation speed at 176.0 tok/s. Largest context that fits in 20GB VRAM, at 8,192 tokens.
Hardware Benchmarks
Compare 2× RTX 3080 benchmark results and see which models actually earn the best completed quality scores.
GPU-specific specifications for local LLM planning.
Models we recommend for 2× RTX 3080.
Fastest & Largest context
2.6B
Highest token generation speed at 176.0 tok/s. Largest context that fits in 20GB VRAM, at 8,192 tokens.
Best quality
26B
Highest completed quality score at 70.9.
Which models actually run best on the 2× RTX 3080, by task, from community benchmark data.
Completed quality-ranked benchmark results for 2× NVIDIA GeForce RTX 3080.