The data is originally source from (Sun et al,2021). (Liu et al, 2023) processed the data to make it a dataset vis huggingface api with taining/validation/testing splitting
Please cite:
@misc{liu2023enhancing,
title={Enhancing Long-form Text Generation in Mental Health with Task-adaptive Tokenization},
author={Siyang Liu and Naihao Deng and Sahand Sabour and Yilin Jia and Minlie Huang and Rada Mihalcea},
year={2023},
eprint={2310.05317},
archivePrefix={arXiv}… See the full description on the dataset page:
https://huggingface.co/datasets/ssss21212/PsyQA.