Edge0 streams MoE experts off SSD to fit 35B in 3 GB
This story is from 2026-09-12. It is preserved in the archive; the latest stories are on the live feed.
Where does a 35B model go when it only takes 2.9 GB of RAM? I went into Edge0 to find out, and I came out with a different mental model of what a local model costs. What shipped Edge0-AI open-sourced Edge0 the other day: a streaming inference engine under Apache 2.0, plus two preview models built o…
Read the full story at DEV Community — Machine Learning ↗
Timeline · 1 report
- 2026-09-12 05:22 · DEV Community — Machine Learning
Edge0 streams MoE experts off SSD to fit 35B in 3 GB