Learning What to Investigate Next: Meta-Reasoning for Long-Horizon Research Agents
arXiv:2610.02525v1 Announce Type: new Abstract: Long-horizon research agents must decide both how to investigate and what to investigate next as evidence accumulates. This is hard to learn because such decisions are sparse in long execution traces, and their consequences may emerge several investig…
Read the full story at arXiv cs.AI ↗
Timeline · 1 report
- 2026-10-05 04:00 · arXiv cs.AI
Learning What to Investigate Next: Meta-Reasoning for Long-Horizon Research Agents