[P] Pecision models that score every allowed label from the logits: Jebadiah v2.1 (27B, 9B), open weights and self-run benchmark results [P]
I've been building open models that treat a decision as a closed-set scoring problem rather than text generation. The input is structured context plus a typed question with a fixed set of options. The output is a probability for each option, taken from the candidate-label logits, so there's no gene…
Read the full story at r/MachineLearning ↗
Timeline · 1 report
- 2026-10-11 09:33 · r/MachineLearning
[P] Pecision models that score every allowed label from the logits: Jebadiah v2.1 (27B, 9B), open weights and self-run benchmark results [P]
More stories
- An Anthropic AI model sent a false homicide tip to Philadelphia police — TechCrunch AI
- Qwen Image 2.1 Turbo Released -- Hugging Face — r/StableDiffusion
- Microsoft CEO Nadella Calls for ‘Emergency Brake’ on Advanced AI — Bloomberg AI
- Microsoft unveils Microsoft-Decision-1, a fast decision-scoring model trained on Qwen3.5-9B, and says it will soon rebase it on MAI, OpenAI, and other models (Achint Srivastava/Command Line) — Techmeme
- Daily Driving Qwen 3.8 Flash-Next MoE (NVFP4) on RTX 5090 + 128GB RAM — Telemetry & Impressions — r/LocalLLM
- Microsoft's Nadella says AI needs an ‘emergency brake’ that humans control — CNBC Technology
- Philadelphia police receive false homicide tip from Anthropic AI model — The Hill Technology
- How Oracle Uses Codex to Help Business Users Get Answers — OpenAI YouTube
Get the daily brief of stories like this at 6:30 every morning →