How User-AI Mistreatment Occurs and Matters in Conversational Systems?
This story is from 2026-09-15. It is preserved in the archive; the latest stories are on the live feed.
arXiv:2609.13579v1 Announce Type: new Abstract: Safety research often focuses on model-generated harms, but users may also direct hostility, coercion, and adversarial pressure at models. Understanding how and when that occurs is essential for accurately interpreting model behaviour, alignment drift…
Read the full story at arXiv cs.AI ↗
Timeline · 1 report
- 2026-09-15 04:00 · arXiv cs.AI
How User-AI Mistreatment Occurs and Matters in Conversational Systems?