VA-DPO: Valence-Arousal Direct Preference Optimization for Controllable Emotion Generation in Language Models
This story is from 2026-08-24. It is preserved in the archive; the latest stories are on the live feed.
arXiv:2608.20374v1 Announce Type: new Abstract: How precisely can we tell a language model how to feel? Most work on emotional generation answers with a discrete label - happy, angry, sad - which cannot express a target like "mildly downcast but calm." We instead specify the desired affect as a con…
Read the full story at arXiv cs.CL ↗
Timeline · 1 report
- 2026-08-24 04:00 · arXiv cs.CL
VA-DPO: Valence-Arousal Direct Preference Optimization for Controllable Emotion Generation in Language Models