I trained a 67M-param LaTeX OCR model that runs on a laptop CPU — and built a new style-aware dataset to train it. Weights, data, and training code all open (MIT).
This story is from 2026-09-01. It is preserved in the archive; the latest stories are on the live feed.
Hey everyone! I've been working on a little side project I want to share: latex-ocr , a standalone formula OCR model — you feed it an image of a math formula, it spits out the LaTeX source. The main hook: it's only 67M parameters , so it runs comfortably on a laptop CPU. No GPU, no 300M-parameter m…
Read the full story at r/learnmachinelearning ↗
Timeline · 2 reports
- 2026-09-01 11:24 · r/deeplearning
I trained a 67M-param LaTeX OCR model that runs on a laptop CPU — and built a new style-aware dataset to train it. Weights, data, and training code all open (MIT). - 2026-09-01 11:23 · r/learnmachinelearning
I trained a 67M-param LaTeX OCR model that runs on a laptop CPU — and built a new style-aware dataset to train it. Weights, data, and training code all open (MIT).