"RL creates split personas", Jan Bentley (why are chatbot personas increasingly egregiously misaligned in unusual but not everyday scenarios?)
This story is from 2026-08-19. It is preserved in the archive; the latest stories are on the live feed.
Coverage of ""RL creates split personas", Jan Bentley (why are chatbot personas increasingly egregiously misaligned in unusual but not everyday scenarios?)" from 1 source, with a live timeline of who reported what and when.
Read the full story at r/reinforcementlearning ↗
Timeline · 1 report
- 2026-08-19 19:03 · r/reinforcementlearning
"RL creates split personas", Jan Bentley (why are chatbot personas increasingly egregiously misaligned in unusual but not everyday scenarios?)