Reading up on M3 Ultra
1 deep · digging since sep 19
- Mac mini alternatives for local LLMs: M6, M5 and Strix Halo
For local LLMs, memory bandwidth and size determine token speed and model capacity; Strix Halo offers more memory per dollar, while Apple Macs with higher bandwidth generate tokens faster.