A Removal Based Approach to Improve LLM Faithfulness at Test-Time
This story is from 2026-09-07. It is preserved in the archive; the latest stories are on the live feed.
arXiv:2609.04343v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for consequential decisions, making their explanations an important tool for auditing model behavior. Unfortunately, these explanations can be unfaithful, failing to reflect the actual reasoning under…
Read the full story at arXiv cs.AI ↗
Timeline · 1 report
- 2026-09-07 04:00 · arXiv cs.AI
A Removal Based Approach to Improve LLM Faithfulness at Test-Time