Distilling a decision API into a head on a frozen encoder with a calibrated fallback loop: numbers, the cache trick, and where it fails
Project write-up with numbers rather than claims. stuntd records typed decisions (choice / score / yes-no) an app asks an LLM for and trains, per decision site, the head Laya ships (type embedding, two transformer layers, a scorer over option markers) on a frozen 400M ModernBERT encoder. Temperatur…
Read the full story at r/deeplearning ↗
Timeline · 1 report
- 2026-09-23 08:38 · r/deeplearning
Distilling a decision API into a head on a frozen encoder with a calibrated fallback loop: numbers, the cache trick, and where it fails
More stories
- GPT-6 Sol and Luna now available on AI Gateway — Vercel Blog
- Alibaba unveils new AI chip to challenge NVIDIA, plans Qwen models with up to 10 trillion parameters — Mint AI
- No Shirt, No Shoes, No Service: Amazon Blocks Meta’s Muse AI From Shopping — CNET AI
- Anthropic launches Claude Opus 5.5 with stricter safeguards for cybersecurity — The Verge AI
- British Columbia Sues OpenAI, Alleging ChatGPT Aided Mass School Shooting — Wall Street Journal Technology
- Moonshot’s Kimi K3 lands on Amazon in key test for Chinese open-source AI revenue — South China Morning Post Tech
- How Benchling secured multi-tenant AI agents with Amazon Bedrock AgentCore — AWS Machine Learning Blog
- Meet the Data Agent in ChatGPT Work — OpenAI YouTube
Get the daily brief of stories like this at 6:30 every morning →