AINewsnow

Running a 180B-Parameter LLM on a Laptop Without a GPU: VIDRAFT's POCKET-Darwin-180B

Running a 180B-Parameter LLM on a Laptop Without a GPU: VIDRAFT's POCKET-Darwin-180B TL;DR: VIDRAFT has released POCKET-Darwin-180B, a 4-bit GGUF-quantized, llama.cpp-compatible build of their Darwin-180B-RSI frontier model that runs on consumer hardware — including CPU-only laptops and mini PCs —…

Read the full story at DEV Community — Machine Learning ↗

Timeline · 1 report

  1. 2026-10-02 23:01 · DEV Community — Machine Learning
    Running a 180B-Parameter LLM on a Laptop Without a GPU: VIDRAFT's POCKET-Darwin-180B

More stories

  1. Benchmarks: Best engine for Qwen 3.8-Flash-Next on Strix Halo — r/LocalLLM
  2. add GLM-5.3-Flash (GLM5-Next) support by timkhronos · Pull Request #27773 · ggml-org/llama.cpp — r/LocalLLaMA
  3. Qwen3.8 flash next ISTA-DASLab GGUF 50t/s TG and 1500t/s PP with 12GB VRAM and 64GB RAM Laptop on 'Strata' engine — r/LocalLLaMA
  4. Sharing my Qwen3.8-27B at 8-bit on 2x RTX 3090 with vLLM: 115 tok/s decode, ~1,780 tok/s prefill, 262K context (NVLink + DFlash2, full recipe and A/B numbers) — r/LocalLLM
  5. Browser FPS with 3D models, textures and SFX generated locally on one GPU, plus a local Qwen 27B for part of the code: my pipeline and what failed — r/LocalLLM
  6. New in llama.cpp: Decision Models — r/LocalLLaMA
  7. Inside-Out AI: Rebuilding Airbnb Behind the Scenes and Across the Guest Experience — Latent Space
  8. Small models are actually quite capable when the input is structured right... — r/LocalLLM

Get the daily brief of stories like this at 6:30 every morning →