Will it run? / Llama 3.1 70B / Ryzen AI Max+ 395 mini PC (128GB)

Can the Ryzen AI Max+ 395 mini PC (128GB) run Llama 3.1 70B?

Too slow to enjoy Yes, at about 4.2 tokens/sec (estimated). It technically runs, but slowly enough to be frustrating.

4.2estimated tokens/sec
44.5 GBmemory needed
96 GBavailable to the model
256 GB/smemory bandwidth
ollama run llama3.1:70b

Check the Ryzen AI Max+ 395 mini PC (128GB) price →

How we got this number

Each generated token reads the whole model from memory once: 43 GB. The Ryzen AI Max+ 395 mini PC (128GB) moves 256 GB/s, so the ceiling is 6.0 tokens/sec. Real machines reach a fraction of that ceiling; we use 70%, an assumption until we get measured runs for this kind of machine. Send us your real numbers and we'll replace the estimate.

Faster machines for Llama 3.1 70B

Mac Studio M5 Ultra (96GB)23 tok/s$5,499Check price →
Mac Studio M5 Ultra (256GB)23 tok/s$9,499Check price →
Mac Studio M5 Max (40-core GPU, 64GB)11 tok/s$3,799Check price →
Mac mini M5 Pro (64GB)5.7 tok/sCheck price →

Other models on the Ryzen AI Max+ 395 mini PC (128GB)

Qwen 3 30B (A3B)Fast52 tok/s
Qwen3 Coder 30B (A3B)Fast52 tok/s
Qwen 3-VL 30B (A3B)Fast50 tok/s
Qwen3.5 35B (A3B)Fast47 tok/s
Qwen3.6 35B (A3B)Fast47 tok/s
gpt-oss 20BFast44 tok/s