I trained a tiny GPT (800K params) on arithmetic and ran 20 controlled experiments to find out what actually helps. Full results inside.
As a learning project, I picked arithmetic expression computation — small enough for my Mac, but it let me run controlled experiments on almost every LLM training method I'd read about: chain-of-thought formats, data scaling, capacity scaling, RoPE, MoE, RLVR, DPO, best-of-N sampling. Every experim…
Read the full story at r/learnmachinelearning ↗
Timeline · 1 report
- 2026-09-21 09:15 · r/learnmachinelearning
I trained a tiny GPT (800K params) on arithmetic and ran 20 controlled experiments to find out what actually helps. Full results inside.