VA-DPO: Valence-Arousal Direct Preference Optimization for controllable emotion generation in language models
Read the original at arxiv.org→arXiv:2608.20374v1 Announce Type: new Abstract: How precisely can we tell a language model how to feel? Most work on emotional generation answers with a discrete label - happy, angry, sad - which cannot express a...
Original headline: "VA-DPO: Valence-Arousal Direct Preference Optimization for Controllable Emotion Generation in Language Models"
Coverage timeline
- Aug 24, 04:00 UTC arXiv cs.CL lead source VA-DPO: Valence-Arousal Direct Preference Optimization for Controllable Emotion Generation in Language Models