Huge upgrade for local AI: Apple introduces M6 and M5 Ultra for a "big leap in performance and AI compute." Capable of running even quantized 70B-class models locally
The new M6 model offers (tl;dr):
• Up to 4.8x faster LLM prompt processing than M4 • 42% more memory bandwidth at 170GB/s • 12-core CPU and GPU • Dual 16-core Neural Engine • Up to 32GB unified memory
The M5 Pro version goes up to 64GB RAM and 307GB/s, making it capable of running even quantized 70B-class models locally.
The catch: 32GB limits the regular M6 to smaller models. It should be excellent for 7B–14B models and usable for quantized 30B models, but serious local AI users should choose the 64GB M5 Pro. Anyways, really really cool release!!