LongEvoRoleBench is a unified benchmark for long-horizon, evolution-aware role-playing.
It standardizes 8 existing character-dialogue corpora into a common next-utterance protocol:
4 long-dialogue corpora test cross-episode character-state evolution, and 4 short-dialogue
corpora provide within-scene state-tracking checks under the same evaluation format.
This repository ships both the raw source corpora and the fully processed evaluation splits.
The… See the full description on the dataset page:
https://huggingface.co/datasets/IAAR-Shanghai/LongEvoRoleBench.