AINewsnow

Ported a 1.3B video world model to Apple Silicon: 6.8s of 832x464 on an M4 Pro, 16 GB peak, no CUDA

This story is from 2026-09-13. It is preserved in the archive; the latest stories are on the live feed.

I spent a couple of days getting LingBot-World-V2 (Robbyant's interactive world model, 1.3B causal-fast variant) running on a single Mac via PyTorch MPS. Upstream assumes 8x CUDA GPUs with torchrun , FSDP and flash_attn . None of that exists on a Mac. Video of the output: https://youtu.be/X8TZXjQSv…

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-09-13 04:20 · r/LocalLLM
    Ported a 1.3B video world model to Apple Silicon: 6.8s of 832x464 on an M4 Pro, 16 GB peak, no CUDA

More stories

  1. A closer look at the upcoming Siri AI-powered home hub, a key pillar of Apple's strategy for the home; sources: Apple started cutting Fitness+ staff (Mark Gurman/Bloomberg) — Techmeme
  2. Gemini Joins the Hacker Club — Wall Street Journal Technology
  3. Apple’s Home AI Hub Details; Apple Fitness+ Layoffs and iPhone Duo Apple Pencil — Bloomberg AI
  4. ComfyUI on Apple Silicon: no MLX, no fp8, 600-second kernel builds. So I built my own launcher — a personal project I'm sharing in case it helps someone. — r/comfyui
  5. [Release] Nirvana Code: A single-binary Rust LLM engine built from the metal up for Apple Silicon (Metal 3, Persistent Prefix Cache, Speculative Decoding, Dual GGUF + MLX) — r/LocalLLM
  6. He’s the Face of AI Doomsday Fears — Wall Street Journal Technology
  7. Week in review: OpenAI ships managed Agents API, Apple's new Siri reportedly runs on Gemini, and three vendors add agent spend controls — r/artificial
  8. Running ACE-Step 1.5 and YuE2-3B on one GPU behind a single local UI (CUDA + Apple Silicon): notes from building it — r/LocalLLM

Get the daily brief of stories like this at 6:30 every morning →