Skip to content
Back to home

Blog.

Explorations at the frontier of Local AI.

14
Distributed training

Local AI is not just about inference

A deep dive into training across slow networks: how DiLoCo works, why checkpoint averaging helps, and how OverlapSPARTA puts more compute to work.

Written in January 2026
Clustering

NVIDIA DGX Spark™ + Apple Mac Studio = 4x Faster LLM Inference with EXO 1.0

How Exo combines DGX Spark’s compute with Mac Studio’s memory bandwidth, splitting prompt processing and token generation across the two devices.

The complete collection01—12