DocHop: Benchmarking Out-of-domain Multi-hop Reasoning in Information-Dense Documents
This story is from 2026-09-03. It is preserved in the archive; the latest stories are on the live feed.
arXiv:2609.02059v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have achieved strong performance on structured visual understanding tasks such as chart and document question answering. However, existing benchmarks typically evaluate these domains in isolation, leaving under…
Read the full story at arXiv cs.AI ↗
Timeline · 1 report
- 2026-09-03 04:00 · arXiv cs.AI
DocHop: Benchmarking Out-of-domain Multi-hop Reasoning in Information-Dense Documents