Hardware Benchmarks

Best Local LLM Performance for Apple M1

Compare Apple M1 benchmark results and see which models actually earn the best completed quality scores.

ShareRedditX

Hardware Overview

GPU-specific specifications for local LLM planning.

Apple M1
VRAM
16GB

Recommended Models

Models we recommend for Apple M1.

Fastest

ling-tiny

Unknown

3774429.8 tok/s

Highest token generation speed at 3774429.8 tok/s.

Best quality & Largest context

ling-oq8e

Unknown

33.8 tok/sQuality 50.0Context 8,192 tokens

Highest completed quality score at 50.0. Largest context that fits in 16GB VRAM, at 8,192 tokens.

Frequently asked questions

Which models actually run best on the Apple M1, by task, from community benchmark data.

What is the best local model for coding on the Apple M1?
On the Apple M1, ling-tiny ranks first for coding in our benchmarks (14.1/100 at ~1887238.2 tok/s).
What is the best local model for agentic workflows on the Apple M1?
On the Apple M1, ling-oq8e ranks first for agentic workflows in our benchmarks (72.6/100 at ~33.8 tok/s).

Benchmark Results

Completed quality-ranked benchmark results for Apple M1.