Fastest
mtplx-qwen38-27b-oq4e-fp16-mtp
27B
486.9 tok/s
Highest token generation speed at 486.9 tok/s.
Hardware Benchmarks
Compare Apple M5 Pro benchmark results and see which models actually earn the best completed quality scores.
GPU-specific specifications for local LLM planning.
Models we recommend for Apple M5 Pro.
Fastest
27B
Highest token generation speed at 486.9 tok/s.
Best quality
27B
Highest completed quality score at 85.1.
Largest context
30.5B
Largest context that fits in 64GB VRAM, at 65,536 tokens.
Which models actually run best on the Apple M5 Pro, by task, from community benchmark data.
Completed quality-ranked benchmark results for Apple M5 Pro.