arXiv paper proposes emotion regulation framework for empathetic speech dialogue in audio-language models
A new arXiv preprint introduces ER-EDF, a framework that draws on psychological theories of emotion perception and regulation to guide empathetic responses in spoken dialogue systems built on large audio-language models. The work aims to improve how such systems both recognize a speaker's emotional state and regulate their own generated reply. It is a research contribution and has not been presented as a product or model release.