Fastest & Best quality & Largest context
ornith-1.0-35b
35B
47.8 tok/sQuality 82.6Context 8,192 tokens
Highest token generation speed at 47.8 tok/s. Highest completed quality score at 82.6. Largest context that fits in 96GB VRAM, at 8,192 tokens.
Hardware Benchmarks
Compare Apple M2 Max benchmark results and see which models actually earn the best completed quality scores.
GPU-specific specifications for local LLM planning.
Models we recommend for Apple M2 Max.
Fastest & Best quality & Largest context
35B
Highest token generation speed at 47.8 tok/s. Highest completed quality score at 82.6. Largest context that fits in 96GB VRAM, at 8,192 tokens.
Which models actually run best on the Apple M2 Max, by task, from community benchmark data.
Completed quality-ranked benchmark results for Apple M2 Max.