Pooled RAM across an old laptop, a Windows PC, and a Mac to run a 13B model - source-available, would love eyes on it
This story is from 2026-09-07. It is preserved in the archive; the latest stories are on the live feed.
https://i.redd.it/ecd8gk9ox0oh1.gif Been running a heterogeneous home cluster for a while — an old Acer laptop (12GB, CPU-only) as the primary API server, with a Windows box (RTX 3060 CUDA) and a Mac Mini (Metal) lending capacity over the network. I wrote the orchestration on top of llama.cpp's `gg…
Read the full story at r/LocalLLM ↗