AINewsnow

NIRNAY: a 450M open model that beats Jev

NIRNAY phase_b scores 0.8792 on the full Banking77 test set (3,080 cases). Jev 1.13.0 scores 0.803 on the same cases (jevbench.xyz recorded run). Ours runs local in ~200ms, no API key, Apache-2.0. The caveat up front, because it matters: we fine-tuned on the train split, Jev answered zero-shot. Sam…

Read the full story at DEV Community — Machine Learning ↗

Timeline · 1 report

  1. 2026-10-01 06:59 · DEV Community — Machine Learning
    NIRNAY: a 450M open model that beats Jev

More stories

  1. Ollama now supports Jev-style decision models — Ollama Blog
  2. I rebuilt a Jev-style classifier on Qwen3.5-4B: shared-prefix tree, open weights, fine-tunable, ~140 ms on one H100 — r/learnmachinelearning
  3. Trained locally: ultra-fast 0.8B/2B System 1 decision models that match Jev on benchmarks and Doom, ~30 ms per decision (open weights) — r/LocalLLaMA
  4. New frontier AI models, TypeSafe’s Jev AI, & NASA’s IBM collab — Mixture of Experts (IBM)
  5. OpenAI's Jev clone could help the frontier lab stop its swarming agents — TechCrunch AI
  6. LLM or JEV? Why not both? - introducing a hybrid Gemma4 approach — r/LocalLLM
  7. Laya/JEV play Magic the Gathering — r/LocalLLM
  8. Mercury released Mercury Decide, I benchmarked it. — r/ArtificialInteligence

Get the daily brief of stories like this at 6:30 every morning →