Blog.
Explorations at the frontier of Local AI.
14 ↓
Latest
Local AI is not just about inference
A deep dive into training across slow networks: how DiLoCo works, why checkpoint averaging helps, and how OverlapSPARTA puts more compute to work.
Written in January 2026
NVIDIA DGX Spark™ + Apple Mac Studio = 4x Faster LLM Inference with EXO 1.0
How Exo combines DGX Spark’s compute with Mac Studio’s memory bandwidth, splitting prompt processing and token generation across the two devices.
The complete collection01—12


