Will it run? / Code Llama 13B / GEEKOM A6 (Ryzen 7 6800H, 32GB)

Can the GEEKOM A6 (Ryzen 7 6800H, 32GB) run Code Llama 13B?

Usable, but slow Yes, at about 6.2 tokens/sec (estimated). It works, but you'll wait on longer answers.

6.2estimated tokens/sec
8.9 GBmemory needed
27.2 GBavailable to the model
76.8 GB/smemory bandwidth
ollama run codellama:13b

Check the GEEKOM A6 (Ryzen 7 6800H, 32GB) price →

How we got this number

Each generated token reads the whole model from memory once: 7.4 GB. The GEEKOM A6 (Ryzen 7 6800H, 32GB) moves 76.8 GB/s, so the ceiling is 10 tokens/sec. Real machines reach a fraction of that ceiling; we use 60%, an assumption until we get measured runs for this kind of machine. Send us your real numbers and we'll replace the estimate.

Faster machines for Code Llama 13B

Mac Studio M5 Ultra (96GB)120+ tok/s$5,499Check price →
Mac Studio M5 Ultra (256GB)120+ tok/s$9,499Check price →
Mac Studio M5 Max (40-core GPU, 64GB)66 tok/s$3,799Check price →
Mac mini M5 Pro (24GB)33 tok/s$1,699Check price →

Other models on the GEEKOM A6 (Ryzen 7 6800H, 32GB)

lfm 2.5 8B (A1B)Fast38 tok/s
minicpm v4.6 1BFast29 tok/s
Gemma 2 2BFast29 tok/s
codeGemma 2BFast29 tok/s
Gemma 2BFast27 tok/s
smollm 2 1.7BFast26 tok/s