Hardware Benchmarks

Best Local LLM Performance for Tesla P100 Pcie 16gb

Compare Tesla P100 Pcie 16gb benchmark results and see which models actually earn the best completed quality scores.

ShareRedditX

Hardware Overview

GPU-specific specifications for local LLM planning.

Tesla P100-PCIE-16GB
VRAM
16GB

Recommended Models

Models we recommend for Tesla P100 Pcie 16gb.

Fastest

Tiel-Coder-35B-A3B-UD-IQ3_XXS

35B

57.7 tok/s

Highest token generation speed at 57.7 tok/s.

Best quality

Cyber-Tiel-Coder-35B-A3B-MTP-UD-Q4_K_XL

35B

38.0 tok/sQuality 85.5

Highest completed quality score at 85.5.

Largest context

Dirk-Qwen3.8-27B-UD-Q4_K_XL

27B

6.2 tok/sContext 65.536 tokens

Largest context that fits in 16GB VRAM, at 65.536 tokens.

Frequently asked questions

Which models actually run best on the Tesla P100 Pcie 16gb, by task, from community benchmark data.

What is the best local model for coding on the Tesla P100 Pcie 16gb?
On the Tesla P100 Pcie 16gb, Dirk-Qwen3.8-27B-UD-Q4_K_XL ranks first for coding in our benchmarks (78.6/100 at ~6.2 tok/s).
What is the best local model for agentic workflows on the Tesla P100 Pcie 16gb?
On the Tesla P100 Pcie 16gb, Cyber-Tiel-Coder-35B-A3B-MTP-UD-Q4_K_XL ranks first for agentic workflows in our benchmarks (83.7/100 at ~38.3 tok/s).

Benchmark Results

Completed quality-ranked benchmark results for Tesla P100-PCIE-16GB.