PowerBench: Measuring Language Model Bias in Power-shifting Requests
arXiv:2610.02303v1 Announce Type: new Abstract: Language models increasingly assist people with power-related requests, so systematic differences in whom they help could shift the distribution of power at scale, or be exploited by users who learn which identities are refused less. We introduce Powe…
Read the full story at arXiv cs.LG ↗
Timeline · 1 report
- 2026-10-05 04:00 · arXiv cs.LG
PowerBench: Measuring Language Model Bias in Power-shifting Requests