Changelog / release

0.13.4 — Qualified Qwen models go faster on their own, and benchmarks stay on your Mac

Released 2026-09-03 · full changelog · GitHub releases

Upgrade: pip install -U rapid-mlx  ·  brew upgrade rapid-mlx  ·  or grab the desktop app.

Four qualified Qwen artifacts now pick their validated speculative-decoding preset without being asked, so concurrent work finishes sooner with no flag to remember. Video generation gets a real workspace in the Mac app, the Community Benchmark keeps every result local until you explicitly share one, and a model that will not fit now says so before it tries to load.