Stable Policy Learning
This story is from 2026-09-18. It is preserved in the archive; the latest stories are on the live feed.
arXiv:2609.19418v1 Announce Type: cross Abstract: In evidence-based policymaking, typically one experimental sample is observed, then a learned policy recommendation is implemented at scale. Policies learned from the experimental data can perform well in expected welfare, yet random sampling in the…
Read the full story at arXiv stat.ML ↗
Timeline · 1 report
- 2026-09-18 04:00 · arXiv stat.ML
Stable Policy Learning