Models · family

Image generation

12 aliases · FLUX.2 klein · Z-Image-Turbo · Qwen-Image · FLUX.1-schnell · HiDream-O1 · Stable Diffusion XL / 3.5 · Bonsai Image. Local image generation through an OpenAI-compatible /v1/images API — the same server the desktop app's Images tab uses.

Pick one

One command per model, smallest memory floor first. The floor is the alias's minimum-memory setting; the size is what rapid-mlx models --modality image-gen reports. Install the extra once with pip install rapid-mlx==0.16.0 (Python 3.11+); weights download on first use.

rapid-mlx serves POST /v1/images/generations (and POST /v1/images/edits) in the OpenAI Images shape, rendering through mflux and pinned native MLX runtimes entirely on your Mac. One image model is resident at a time — same single-worker discipline as the video lane. In the desktop app, the Images tab drives the same endpoints: pick a model, prompt, refine, and every render lands in a filmstrip you can step back through.

family
Image generation (mflux + native MLX)
aliases
12
install
pip install rapid-mlx==0.16.0
OpenAI base URL
http://localhost:8000/v1

Usage

pip install rapid-mlx==0.16.0
rapid-mlx serve z-image-turbo

curl http://localhost:8000/v1/images/generations \
  -H 'Content-Type: application/json' \
  -d '{"model":"z-image-turbo","prompt":"a lighthouse at dusk, oil painting","size":"1024x1024"}'

Family pages

Z-Image, FLUX and Qwen-Image have their own pages with usage and download details:

Notes & caveats

Frequently asked questions

Can I generate images locally on a Mac?

Yes — POST /v1/images/generations renders entirely on Apple Silicon. No cloud, no per-image cost. The desktop Images tab uses the same server.

Which image models run on Apple Silicon?

Twelve aliases, from qwen-image-2.1 and flux2-klein-4b (8–12 GB Macs) up to qwen-image-edit (96 GB). flux2-klein-4b is the fast default.

Where next