Safe Meta-Policy Design with Risk Control
arXiv:2610.10393v1 Announce Type: new Abstract: Models can be retrained as new data arrive, but deploying every new version risks replacing a good policy with a worse one. We study how to plan policy updates (i.e., meta-policy) before future candidates are trained, balancing the benefits of improve…
Read the full story at arXiv stat.ML ↗
Timeline · 1 report
- 2026-10-08 04:00 · arXiv stat.ML
Safe Meta-Policy Design with Risk Control