MindChat-R0: A Large Language Model for Emotionally Supportive Dialogue through Reinforcement Learning
Abstract
MindChat-R0 is a Chinese emotionally supportive dialogue model trained through reinforcement learning with structured emotional reasoning. The work introduces an Empathetic Chain-of-Thought framework and trains the model using reward signals designed to improve empathy, fluency, coherence, and supportive responses for people facing mental health challenges.
Type
Publication
Companion of the 2025 ACM International Joint Conference on Pervasive and Ubiquitous Computing (UbiComp Companion 2025), pp. 1209-1216
Recognition: 10th Mental Health Workshop Best Paper Honorable Mention