Models · family

Hunyuan 3 (Hy3)

1 MLX alias · Tencent Hunyuan 3 · 295B MoE · preview · M3-Ultra-class (256 GB+) · hy_v3 tool + reasoning parser.

Hunyuan 3 (Hy3) is Tencent's 295B mixture-of-experts flagship, shipped in MLX form as a preview. It uses a dedicated hy_v3 tool-call parser and hy_v3 reasoning parser, and honors a reasoning_effort override on /v1/responses and /v1/chat/completions. This is an Ultra-class model — the weights want an M3 Ultra with 256 GB+ of unified memory, so it is not something a 32 or 64 GB Mac can hold.

Not R2-mirrored. Hy3 is not on the models.rapidmlx.com RAM picker or the R2 catalog — it pulls from Hugging Face on first use via its alias. Run rapid-mlx pull hy3-preview-4bit (or just serve) and the weights stream from mlx-community/Hy3-preview-4bit.

family
Hunyuan
aliases
1
status
preview
min RAM
256 GB (M3 Ultra-class)
install
rapid-mlx serve <alias>
OpenAI base URL
http://localhost:8000/v1

Download

Every alias on this page downloads with one command — the pull buttons in the tables below copy it. These checkpoints pull from Hugging Face directly (not mirrored on our CDN). Weights land in the standard Hugging Face cache, and rapid-mlx serve pulls automatically on first use. Live mirror status →

Hunyuan 3 295B · 1 alias · preview

295B MoE in 4-bit MLX. Dedicated hy_v3 tool-call parser and reasoning parser, with a reasoning_effort override. Ultra-class: plan for 256 GB+ of unified memory.

tool parser: hy_v3 · reasoning parser: hy_v3

aliashf repotool parserreasoningflagsget it
hy3-preview-4bitmlx-community/Hy3-preview-4bithy_v3hy_v3moe · preview · 256 GB · no-specHF

Pull & serve

rapid-mlx pull hy3-preview-4bit
rapid-mlx serve hy3-preview-4bit

Notes & caveats

Where next