I trained a 1.5B model to be less overconfident instead of pretending uncertainty doesn't exist
I've been working on a small independent research project called Vera (Verifiable Epistemic Reward Alignment) . The idea is pretty simple: LLMs can produce answers with a very high-confidence tone even when the underlying claim is uncertain, controversial, or simply false. Instead of trying to make…
Read the full story at r/huggingface ↗
Timeline · 1 report
- 2026-10-02 01:19 · r/huggingface
I trained a 1.5B model to be less overconfident instead of pretending uncertainty doesn't exist