Hardware Benchmarks

Best Local LLM Performance for RTX 3080 + RTX 3080 Ti

Compare RTX 3080 + RTX 3080 Ti benchmark results and see which models actually earn the best completed quality scores.

ShareRedditX

Hardware Overview

GPU-specific specifications for local LLM planning.

NVIDIA GeForce RTX 3080 + NVIDIA GeForce RTX 3080 Ti
VRAM
32GB
Architecture
Multi-GPU

Recommended Models

Models we recommend for RTX 3080 + RTX 3080 Ti.

Fastest

LiquidAI/LFM2.5-350M-GGUF:F16

Unknown

665.0 tok/s

Highest token generation speed at 665.0 tok/s.

Best quality

unsloth/Qwen3.8-27B-GGUF:UD-Q6_K_XL

27B

47.1 tok/sQuality 89.6

Highest completed quality score at 89.6.

Frequently asked questions

Which models actually run best on the RTX 3080 + RTX 3080 Ti, by task, from community benchmark data.

What is the best local model for coding on the RTX 3080 + RTX 3080 Ti?
On the RTX 3080 + RTX 3080 Ti, Qwen3.8-27B-SC_6.00bpw_H6 ranks first for coding in our benchmarks (83.6/100 at ~41.8 tok/s). unsloth/Muse-Glimmer-30B-GGUF:UD-Q6_K_XL is faster (~52.3 tok/s) and still scores well (81.0/100) — a good pick if you'd rather trade a little quality for speed.
What is the best local model for agentic workflows on the RTX 3080 + RTX 3080 Ti?
On the RTX 3080 + RTX 3080 Ti, unsloth/Qwen3.8-27B-GGUF:UD-Q6_K_XL ranks first for agentic workflows in our benchmarks (92.0/100 at ~47.1 tok/s). unsloth/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-GGUF:MXFP4_MOE is faster (~162.1 tok/s) and still scores well (91.1/100) — a good pick if you'd rather trade a little quality for speed.

Benchmark Results

Completed quality-ranked benchmark results for NVIDIA GeForce RTX 3080 + NVIDIA GeForce RTX 3080 Ti.