Benchmarks · community-measured

Apple Silicon LLM benchmarks

Measured local-inference performance, contributed by rapid-mlx users: 22 model-and-chip combinations from 44 submissions across 6 machines (updated 2026-08-25). Explore and filter on the interactive leaderboard, or jump straight to a measurement below. Every number is reproducible: rapid-mlx bench <model> --tier speed --submit.

By model

bonsai-1.7b-2bit

bonsai-27b-2bit

gemma-4-26b-4bit

gpt-oss-20b-mxfp4-q8

qwen2.5-14b-4bit

qwen3.5-27b-4bit

qwen3.5-4b-4bit

qwen3.5-9b-4bit

qwen3.5-9b-8bit

qwen3.6-27b-4bit

qwen3.6-35b-4bit

qwen3.6-35b-optiq-4bit

qwen3.8-27b-4bit

qwen3.8-27b-mixed-3.5bpw

By chip

Apple M2

Apple M2 Pro

Apple M4

Apple M4 Max

Apple M5 Max

Apple M5 Pro

Add your machine

$ curl -fsSL https://rapidmlx.com/install.sh | bash
$ rapid-mlx bench qwen3.5-9b-4bit --tier speed --submit

Raw corpus: /api/benchmarks · Submissions are anonymous and retractable — see the leaderboard for details.