Satisficing Regret Minimization in Bandits: Constant Rate and Light-Tailed Distribution
This story is from 2026-09-14. It is preserved in the archive; the latest stories are on the live feed.
arXiv:2406.06802v4 Announce Type: replace Abstract: Motivated by the concept of satisficing in decision-making, we consider the problem of satisficing regret minimization in bandit optimization. In this setting, the learner aims at selecting satisficing arms (arms with mean reward exceeding a certa…
Read the full story at arXiv stat.ML ↗
Timeline · 1 report
- 2026-09-14 04:00 · arXiv stat.ML
Satisficing Regret Minimization in Bandits: Constant Rate and Light-Tailed Distribution