Hardware Benchmarks

Best Local LLM Performance for 2× RTX 3090 Ti

Compare 2× RTX 3090 Ti benchmark results and see which models actually earn the best completed quality scores.

ShareRedditX

Hardware Overview

GPU-specific specifications for local LLM planning.

2× NVIDIA GeForce RTX 3090 Ti
VRAM
46GB
Architecture
Multi-GPU

Recommended Models

Models we recommend for 2× RTX 3090 Ti.

Fastest & Best quality & Largest context

funera_main:latest

24.0B

52.7 tok/sQuality 46.1Context 16,384 tokens

Highest token generation speed at 52.7 tok/s. Highest completed quality score at 46.1. Largest context that fits in 46GB VRAM, at 16,384 tokens.

Frequently asked questions

Which models actually run best on the 2× RTX 3090 Ti, by task, from community benchmark data.

What is the best local model for coding on the 2× RTX 3090 Ti?
On the 2× RTX 3090 Ti, funera_main:latest ranks first for coding in our benchmarks (15.2/100 at ~52.7 tok/s).
What is the best local model for agentic workflows on the 2× RTX 3090 Ti?
On the 2× RTX 3090 Ti, funera_ultim2.0:latest ranks first for agentic workflows in our benchmarks (52.7/100 at ~52.6 tok/s).

Benchmark Results

Completed quality-ranked benchmark results for 2× NVIDIA GeForce RTX 3090 Ti.