Evaluating lexical vs neural semantic entropy across 1.5B to 120B models
This story is from 2026-09-06. It is preserved in the archive; the latest stories are on the live feed.
We recently released a benchmark looking at uncertainty estimation across model sizes, comparing Farquhar et al.'s Semantic Entropy (Nature, 2024) against a normalized exact-match entropy metric (R_sc). Preprint: https://zenodo.org/records/22233648 Code: https://github.com/Adarshent/Spnda The issue…
Read the full story at r/deeplearning ↗
Timeline · 1 report
- 2026-09-06 11:03 · r/deeplearning
Evaluating lexical vs neural semantic entropy across 1.5B to 120B models