More Programs or More Rolls? Separating Coverage from Specialization in LLM Harnesses
arXiv:2609.35873v1 Announce Type: new Abstract: Automated generation of LLM harnesses promises to improve inference through task specialization. Yet additional answer coverage can arise from repeated execution of the same program, making specialization difficult to identify. We introduce a controll…
Read the full story at arXiv cs.AI ↗
Timeline · 1 report
- 2026-09-30 04:00 · arXiv cs.AI
More Programs or More Rolls? Separating Coverage from Specialization in LLM Harnesses