My local model setup on an M4 Pro Mac mini
kept by devopsfolk
Kevin Lewis details his local LLM server on an M4 Pro Mac mini with 48 GB RAM, using Qwen and Gemma models via oMLX and Tailscale for private, cost‑predictable AI.