Hardware Benchmarks

Best Local LLM Performance for RTX 2080 Ti + RTX A2000 12gb

Compare RTX 2080 Ti + RTX A2000 12gb benchmark results and see which models actually earn the best completed quality scores.

ShareRedditX

Hardware Overview

GPU-specific specifications for local LLM planning.

NVIDIA GeForce RTX 2080 Ti + NVIDIA RTX A2000 12GB
VRAM
23GB
Architecture
Multi-GPU

Recommended Models

Models we recommend for RTX 2080 Ti + RTX A2000 12gb.

Fastest

gemma4-26b-uncensored-mtp

26B

55.1 tok/s

Highest token generation speed at 55.1 tok/s.

Best quality & Largest context

qwen38-27b-turbo

27B

29.8 tok/sQuality 78.0Context 8,192 tokens

Highest completed quality score at 78.0. Largest context that fits in 23GB VRAM, at 8,192 tokens.

Frequently asked questions

Which models actually run best on the RTX 2080 Ti + RTX A2000 12gb, by task, from community benchmark data.

What is the best local model for coding on the RTX 2080 Ti + RTX A2000 12gb?
On the RTX 2080 Ti + RTX A2000 12gb, qwen38-27b-turbo ranks first for coding in our benchmarks (72.8/100 at ~29.8 tok/s).
What is the best local model for agentic workflows on the RTX 2080 Ti + RTX A2000 12gb?
On the RTX 2080 Ti + RTX A2000 12gb, qwen38-27b-turbo ranks first for agentic workflows in our benchmarks (85.7/100 at ~29.8 tok/s).

Benchmark Results

Completed quality-ranked benchmark results for NVIDIA GeForce RTX 2080 Ti + NVIDIA RTX A2000 12GB.