Will it run? / Qwen3.5 35B (A3B) / Mac mini M5 Pro (24GB)

Can the Mac mini M5 Pro (24GB) run Qwen3.5 35B (A3B)?

Won't fit No. Qwen3.5 35B (A3B) needs about 25.5 GB of memory, and the Mac mini M5 Pro (24GB) can give a model about 16.1 GB. You'd need a machine with more memory, or a smaller model.

–estimated tokens/sec
25.5 GBmemory needed
16.1 GBavailable to the model
307 GB/smemory bandwidth

Check the Mac mini M5 Pro (24GB) price →

How we got this number

Each generated token reads the active part of the model from memory once: 2.1 GB. The Mac mini M5 Pro (24GB) moves 307 GB/s, so the ceiling is 120+ tokens/sec. Real machines reach a fraction of that ceiling; we use 80%, calibrated against a measured Mac run. Send us your real numbers and we'll replace the estimate.

Faster machines for Qwen3.5 35B (A3B)

Mac Studio M5 Max (40-core GPU, 64GB)120+ tok/s$3,799Check price →
Mac Studio M5 Ultra (96GB)120+ tok/s$5,499Check price →
Mac Studio M5 Ultra (256GB)120+ tok/s$9,499Check price →
Mac mini M5 Pro (64GB)64 tok/sCheck price →

Other models on the Mac mini M5 Pro (24GB)

lfm 2.5 8B (A1B)Fast120+ tok/s
gpt-oss 20BFast60 tok/s
Llama 3.1 8BFast50 tok/s
Qwen 7BFast50 tok/s
dolPhin 3 8BFast50 tok/s
Gemma 7BFast49 tok/s