Ported a 1.3B video world model to Apple Silicon: 6.8s of 832x464 on an M4 Pro, 16 GB peak, no CUDA
This story is from 2026-09-13. It is preserved in the archive; the latest stories are on the live feed.
I spent a couple of days getting LingBot-World-V2 (Robbyant's interactive world model, 1.3B causal-fast variant) running on a single Mac via PyTorch MPS. Upstream assumes 8x CUDA GPUs with torchrun , FSDP and flash_attn . None of that exists on a Mac. Video of the output: https://youtu.be/X8TZXjQSv…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-09-13 04:20 · r/LocalLLM
Ported a 1.3B video world model to Apple Silicon: 6.8s of 832x464 on an M4 Pro, 16 GB peak, no CUDA