A weekend with TensorFold on a MacBook: the engine mattered, the quant did not
This story is from 2026-09-28. It is preserved in the archive; the latest stories are on the live feed.
A 27B dense model on my MacBook decodes at about 26 tokens a second. That is fine for chat and painful for an agent that writes a few thousand tokens per turn across a hundred turns. So when a new engine showed up claiming several times that on the same laptop, with output that is exactly the same…
Read the full story at DEV Community — AI ↗
Timeline · 1 report
- 2026-09-28 17:38 · DEV Community — AI
A weekend with TensorFold on a MacBook: the engine mattered, the quant did not