Hardware Benchmarks

Best Local LLM Performance for RTX 5060 Ti + RTX 5090

Compare RTX 5060 Ti + RTX 5090 benchmark results and see which models actually earn the best completed quality scores.

ShareRedditX

Hardware Overview

GPU-specific specifications for local LLM planning.

NVIDIA GeForce RTX 5060 Ti + NVIDIA GeForce RTX 5090
VRAM
47GB
Architecture
Multi-GPU

Recommended Models

Models we recommend for RTX 5060 Ti + RTX 5090.

Fastest & Best quality & Largest context

swift-1.5-qwen3.8-27b-uncensored

27B

282.2 tok/sQuality 87.9Context 100,000 tokens

Highest token generation speed at 282.2 tok/s. Highest completed quality score at 87.9. Largest context that fits in 47GB VRAM, at 100,000 tokens.

Frequently asked questions

Which models actually run best on the RTX 5060 Ti + RTX 5090, by task, from community benchmark data.

What is the best local model for coding on the RTX 5060 Ti + RTX 5090?
On the RTX 5060 Ti + RTX 5090, swift-1.5-qwen3.8-27b-uncensored ranks first for coding in our benchmarks (77.8/100 at ~275.1 tok/s).
What is the best local model for agentic workflows on the RTX 5060 Ti + RTX 5090?
On the RTX 5060 Ti + RTX 5090, swift-1.5-qwen3.8-27b-uncensored ranks first for agentic workflows in our benchmarks (86.6/100 at ~275.1 tok/s).

Benchmark Results

Completed quality-ranked benchmark results for NVIDIA GeForce RTX 5060 Ti + NVIDIA GeForce RTX 5090.