Fastest
ornith-ai-ornith-1.5-35b-a3b-mtplx-1
35B
51.3 tok/s
Highest token generation speed at 51.3 tok/s.
Hardware Benchmarks
Compare Apple M1 Max benchmark results and see which models actually earn the best completed quality scores.
GPU-specific specifications for local LLM planning.
Models we recommend for Apple M1 Max.
Fastest
35B
Highest token generation speed at 51.3 tok/s.
Best quality & Largest context
27B
Highest completed quality score at 85.7. Largest context that fits in 64GB VRAM, at 262,144 tokens.
Completed quality-ranked benchmark results for Apple M1 Max.