LexAgentHallu: A Hierarchical Benchmark for Profiling Hallucinations in Legal Agents
This story is from 2026-09-11. It is preserved in the archive; the latest stories are on the live feed.
arXiv:2609.09754v1 Announce Type: new Abstract: As large language models are increasingly deployed as tool-augmented legal agents, they introduce agentic hallucinations where tool-call and reasoning errors cascade into fabricated holdings and miscited authority. However, existing legal benchmarks e…
Read the full story at arXiv cs.AI ↗
Timeline · 1 report
- 2026-09-11 04:00 · arXiv cs.AI
LexAgentHallu: A Hierarchical Benchmark for Profiling Hallucinations in Legal Agents