Hardware Benchmarks

Best Local LLM Performance for Apple M1 Max

Compare Apple M1 Max benchmark results and see which models actually earn the best completed quality scores.

ShareRedditX

Hardware Overview

GPU-specific specifications for local LLM planning.

Apple M1 Max
VRAM
64GB

Recommended Models

Models we recommend for Apple M1 Max.

Fastest

ornith-ai-ornith-1.5-35b-a3b-mtplx-1

35B

51.3 tok/s

Highest token generation speed at 51.3 tok/s.

Best quality

mtplx-qwen38-27b-optimized-speed-fp16

27B

17.1 tok/sQuality 85.7

Highest completed quality score at 85.7.

Frequently asked questions

Which models actually run best on the Apple M1 Max, by task, from community benchmark data.

What is the best local model for coding on the Apple M1 Max?
On the Apple M1 Max, mtplx-qwen38-27b-optimized-speed-fp16 ranks first for coding in our benchmarks (84.0/100 at ~17.1 tok/s). ornith-ai-ornith-1.5-35b-a3b-mtplx-1 is faster (~51.3 tok/s) and still scores well (83.9/100) — a good pick if you'd rather trade a little quality for speed.
What is the best local model for agentic workflows on the Apple M1 Max?
On the Apple M1 Max, mtplx-qwen38-27b-optimized-speed-fp16 ranks first for agentic workflows in our benchmarks (85.7/100 at ~17.1 tok/s). ornith-ai-ornith-1.5-35b-a3b-mtplx-1 is faster (~51.3 tok/s) and still scores well (81.6/100) — a good pick if you'd rather trade a little quality for speed.

Benchmark Results

Completed quality-ranked benchmark results for Apple M1 Max.