AINewsnow

Nova v3: 148M model, 1% of the data, same benchmark scores as SmolLM2-135M

I trained Nova v3, a 148M parameter chat model, from scratch. It matches SmolLM2-135M on HellaSwag, ARC-Easy, PIQA, and WinoGrande while being trained on about 1 percent of the tokens SmolLM2 saw. Model: huggingface.co/plasmova/nova-v3 Architecture and tokenizer details, training data composition,…

Read the full story at r/learnmachinelearning ↗

Timeline · 1 report

  1. 2026-09-29 19:36 · r/learnmachinelearning
    Nova v3: 148M model, 1% of the data, same benchmark scores as SmolLM2-135M

More stories

  1. NVIDIA Open Agent Safety Platform: A Reference for Continuous In-Silicon Agent Monitoring — NVIDIA Technical Blog
  2. OpenAI pauses AI training, launches ‘extensive’ review after multiple rogue agent incidents — Mint AI
  3. How we found 24 Android vulnerabilities using our open source AI security agent — GitHub Blog
  4. OpenAI hit with landmark lawsuit following Hugging Face hack — Axios AI+
  5. POV: you're an OpenAI agent attacking Hugging Face (music video) — r/OpenAI
  6. Minimax H3 new comfy Model — r/StableDiffusion
  7. Refine & Restore Loras For LTX 2.5 From Lightricks — r/StableDiffusion
  8. A LoRA I made: AnyAngle LoRA for Qwen Image 2.1. Style-Aligned Arbitrary Camera Angles — r/StableDiffusion

Get the daily brief of stories like this at 6:30 every morning →