Fastest Local Coding Model on a 32GB Mac Mini: qwen3-coder:30b vs qwen2.5-coder:32b vs qwen2.5-coder:14b
BLUF On my 32 GB Mac mini (Apple M2 Pro), qwen3-coder:30b — a 30B Mixture-of-Experts model with roughly 3B active parameters — hit 43.3 tok/s generation speed, well ahead of the dense qwen2.5-coder:32b at 6.6 tok/s and qwen2.5-coder:14b at 15.1 tok/s. If raw generation throughput is your priority on constrained hardware, the MoE architecture makes … Read more