Training a LoRA adapter on Kimi K3 (2.78T params, 1.56TB of weights) on a 2017 laptop with 7.6GB of RAM — 7.4 hours per step, and here's the verification
This story is from 2026-09-10. It is preserved in the archive; the latest stories are on the live feed.
Kimi K3 is a 2.78 T MoE; its 1.56 TB checkpoint sits on a USB hard disk plugged into a 2017 laptop (i7-7700HQ, 7.6 GB of RAM, a 2 GB GTX 1050 that only does the routed-expert matmuls). I am training a LoRA adapter on it out of core: the non-expert weights of one layer at a time, its 896 experts str…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-09-10 15:47 · r/LocalLLM
Training a LoRA adapter on Kimi K3 (2.78T params, 1.56TB of weights) on a 2017 laptop with 7.6GB of RAM — 7.4 hours per step, and here's the verification
More stories
- Introducing Kimi K3 on Amazon Bedrock — AWS Machine Learning Blog
- kimi 💀 — r/ArtificialInteligence
- 2 TB of Cheep Pmem200 Dimms can Run Kimi K3 at tg128 ~ 1 t/s · pp512 5.6558 — r/LocalLLM
- Are we over-engineering AI agent workflows? — r/AI_Agents
- StepFun joins the frontier: a previously non-frontier Chinese lab (StepFun) released a Kimi K3-level model, 3 times cheaper per Artificial Analysis — r/singularity
- A Chinese AI company just connected its model to Wall Street's leading data providers — CNBC Technology
- Anthropic, OpenAI, SpaceXAI, Google sued over call to ‘pace’ AI development — Politico Technology
- Introducing Amazon SageMaker HyperPod Inference Gateway — AWS Machine Learning Blog
Get the daily brief of stories like this at 6:30 every morning →