Fastest & Largest context
deepseek-r1:1.5b
1.8B
54.5 tok/sContext 8.192 tokens
Highest token generation speed at 54.5 tok/s. Largest context that fits in 4GB VRAM, at 8.192 tokens.
Hardware Benchmarks
Compare RTX 2050 benchmark results and see which models actually earn the best completed quality scores.
GPU-specific specifications for local LLM planning.
Models we recommend for RTX 2050.
Fastest & Largest context
1.8B
Highest token generation speed at 54.5 tok/s. Largest context that fits in 4GB VRAM, at 8.192 tokens.
Best quality
4.0B
Highest completed quality score at 61.4.
Which models actually run best on the RTX 2050, by task, from community benchmark data.
Completed quality-ranked benchmark results for NVIDIA GeForce RTX 2050.