Hardware Benchmarks

Best Local LLM Performance for RTX 3080

Compare RTX 3080 benchmark results and see which models actually earn the best completed quality scores.

ShareRedditX

Hardware Overview

GPU-specific specifications for local LLM planning.

NVIDIA GeForce RTX 3080
VRAM
10GB

Recommended Models

Models we recommend for RTX 3080.

Fastest

ornith-1.5-9b

9B

94.0 tok/s

Highest token generation speed at 94.0 tok/s.

Best quality

qwen3_coder_next

Unknown

31.1 tok/sQuality 83.6

Highest completed quality score at 83.6.

Frequently asked questions

Which models actually run best on the RTX 3080, by task, from community benchmark data.

What is the best local model for coding on the RTX 3080?
On the RTX 3080, qwen3_coder_next ranks first for coding in our benchmarks (72.5/100 at ~31.2 tok/s).
What is the best local model for agentic workflows on the RTX 3080?
On the RTX 3080, qwen3_coder_next ranks first for agentic workflows in our benchmarks (87.3/100 at ~31.2 tok/s).

Benchmark Results

Completed quality-ranked benchmark results for NVIDIA GeForce RTX 3080.