Task-grouped multidimensional summarization quality data from MSumBench.
The release contains 2,250 complete artifact rows from 150 source-document groups.
The three prediction targets are faithfulness, completeness, and conciseness.
Each seed has task-grouped train, validation, and test splits. All language-specific summaries and model generations for one group_id remain in exactly one split. The seeds change… See the full description on the dataset page:
https://huggingface.co/datasets/Samsoup/MSumBench.