Edge0 Runs a 35B AI Model on an iPhone With Just 2.9 GiB
This story is from 2026-09-12. It is preserved in the archive; the latest stories are on the live feed.
Edge0 streams a 35B Mixture-of-Experts model from SSD on an iPhone, holding under 3 GB of active RAM while decoding at 15 tokens per second.
Read the full story at AlphaSignal ↗
Timeline · 1 report
- 2026-09-12 19:59 · AlphaSignal
Edge0 Runs a 35B AI Model on an iPhone With Just 2.9 GiB