Fastest
gpt-oss:20b
20.9B
31.1 tok/s
Highest token generation speed at 31.1 tok/s.
Hardware Benchmarks
Compare Apple M1 Pro benchmark results and see which models actually earn the best completed quality scores.
GPU-specific specifications for local LLM planning.
Models we recommend for Apple M1 Pro.
Fastest
20.9B
Highest token generation speed at 31.1 tok/s.
Best quality & Largest context
11.9B
Highest completed quality score at 71.1. Largest context that fits in 32GB VRAM, at 65.536 tokens.
Which models actually run best on the Apple M1 Pro, by task, from community benchmark data.
Completed quality-ranked benchmark results for Apple M1 Pro.