AINewsnow

Training a LoRA adapter on Kimi K3 (2.78T params, 1.56TB of weights) on a 2017 laptop with 7.6GB of RAM — 7.4 hours per step, and here's the verification

This story is from 2026-09-10. It is preserved in the archive; the latest stories are on the live feed.

Kimi K3 is a 2.78 T MoE; its 1.56 TB checkpoint sits on a USB hard disk plugged into a 2017 laptop (i7-7700HQ, 7.6 GB of RAM, a 2 GB GTX 1050 that only does the routed-expert matmuls). I am training a LoRA adapter on it out of core: the non-expert weights of one layer at a time, its 896 experts str…

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-09-10 15:47 · r/LocalLLM
    Training a LoRA adapter on Kimi K3 (2.78T params, 1.56TB of weights) on a 2017 laptop with 7.6GB of RAM — 7.4 hours per step, and here's the verification

More stories

  1. Introducing Kimi K3 on Amazon Bedrock — AWS Machine Learning Blog
  2. kimi 💀 — r/ArtificialInteligence
  3. 2 TB of Cheep Pmem200 Dimms can Run Kimi K3 at tg128 ~ 1 t/s · pp512 5.6558 — r/LocalLLM
  4. Are we over-engineering AI agent workflows? — r/AI_Agents
  5. StepFun joins the frontier: a previously non-frontier Chinese lab (StepFun) released a Kimi K3-level model, 3 times cheaper per Artificial Analysis — r/singularity
  6. A Chinese AI company just connected its model to Wall Street's leading data providers — CNBC Technology
  7. Anthropic, OpenAI, SpaceXAI, Google sued over call to ‘pace’ AI development — Politico Technology
  8. Introducing Amazon SageMaker HyperPod Inference Gateway — AWS Machine Learning Blog

Get the daily brief of stories like this at 6:30 every morning →