I rebuilt a Jev-style classifier on Qwen3.5-4B: shared-prefix tree, open weights, fine-tunable, ~140 ms on one H100
By now everyone has heard of Jev: fast, accurate answers to multiple-choice questions about any text, in one API call. Here's how to get the same thing on your own GPU. SelfJev is a 4B model (Qwen3.5-4B + LoRA) that works like Jev: ⚡ About 140 ms per call on one H100, which is about as fast as Jev'…
Read the full story at r/learnmachinelearning ↗
Timeline · 1 report
- 2026-09-30 07:54 · r/learnmachinelearning
I rebuilt a Jev-style classifier on Qwen3.5-4B: shared-prefix tree, open weights, fine-tunable, ~140 ms on one H100
More stories
- Ollama now supports Jev-style decision models — Ollama Blog
- Trained locally: ultra-fast 0.8B/2B System 1 decision models that match Jev on benchmarks and Doom, ~30 ms per decision (open weights) — r/LocalLLaMA
- OpenDecider: distilling calibrated "System One" decision models from open teachers, evaluated head-to-head against Laya and TypeSafe's Jev — r/machinelearningnews
- New frontier AI models, TypeSafe’s Jev AI, & NASA’s IBM collab — Mixture of Experts (IBM)
- LLM or JEV? Why not both? - introducing a hybrid Gemma4 approach — r/LocalLLM
- How long until OpenAI release their version of Jev? — r/AI_Agents
- The question picks the model: Jev and Claude on three of our tasks — DEV Community — Machine Learning
- autotrust/JEV-27B-VL: first decision model that learned to see without a single image of training — r/LocalLLM
Get the daily brief of stories like this at 6:30 every morning →