The dataset is collected and used in paper Modeling Empathetic Alignment in Conversation. It contains therapeutic Reddit conversations labeled with 9,284 appraisals from both the Target and Observer and 3,262 alignments between the Target and Observer.
The dataset has the following entries:
id
target_id
observer_id
parent_id
subreddit
target_text
observer_text
distress_score
condolence_score
empathy_score
full_text
spans
alignments
int
string
string… See the full description on the dataset page:
https://huggingface.co/datasets/Blablablab/ALOE.