Hardware Benchmarks

Best Local LLM Performance for Apple M5

Compare Apple M5 benchmark results and see which models actually earn the best completed quality scores.

ShareRedditX

Hardware Overview

GPU-specific specifications for local LLM planning.

Apple M5
VRAM
24GB

Recommended Models

Models we recommend for Apple M5.

Fastest & Best quality & Largest context

openai/gpt-oss-20b

20B

45.3 tok/sQuality 74.0Context 8,192 tokens

Highest token generation speed at 45.3 tok/s. Highest completed quality score at 74.0. Largest context that fits in 24GB VRAM, at 8,192 tokens.

Frequently asked questions

Which models actually run best on the Apple M5, by task, from community benchmark data.

What is the best local model for coding on the Apple M5?
On the Apple M5, openai/gpt-oss-20b ranks first for coding in our benchmarks (73.0/100 at ~45.3 tok/s).
What is the best local model for agentic workflows on the Apple M5?
On the Apple M5, openai/gpt-oss-20b ranks first for agentic workflows in our benchmarks (89.4/100 at ~45.3 tok/s).

Benchmark Results

Completed quality-ranked benchmark results for Apple M5.