AINewsnow

I tested 14 "decision models" against local LLMs on 56 real triage tasks. Qwen3.5-9B beat every one of them.

OpenRouter now lists a bunch of "decision models" (Solar Decide, Jev, Microsoft-Decision-1, Clef, Mercury Decide…). You don't chat with them: you send a situation plus multiple-choice questions, and they return a probability for every option. They're pitched for routing, triage and classification.…

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-10-11 11:26 · r/LocalLLM
    I tested 14 "decision models" against local LLMs on 56 real triage tasks. Qwen3.5-9B beat every one of them.

More stories

  1. The maker of non-text AI model Jev valued at $7.5B just weeks after launch — TechCrunch AI
  2. A Practical Guide to OpenAI's New Decisions API — Towards Data Science
  3. [Research] Can a swarm of Jev catch a murderer? — r/LocalLLaMA
  4. Why people didn’t seem to care about jev with vision support! — r/AI_Agents
  5. Created laya : Now Introducing a new 800 Million Param physics-based typed decision model with 73k context and image support — r/LocalLLaMA
  6. [AINews] TypeSafe/Jev at >$100M ARR, $7.5B valuation 3 weeks after launch — Latent Space
  7. Jev vs LLM judges on my agent's evals: 180x cheaper, 2.6 points less accurate — r/AI_Agents
  8. Nace AI Open-Sources Drex 1.5: A 9B Decision Model That Scores Options, Not Text — MarkTechPost

Get the daily brief of stories like this at 6:30 every morning →