Public conversations, original full-dialogue G-Eval, continuity-filtered G-Eval,
source-versus-generated discriminator artifacts, and compact SFT training
analyses for the current PatientAgent comparison.
All response sources use the same source-eligible dialogue IDs from both official MTS-Dialog test splits. Eligibility is defined before generation: the ground-truth source dialogue must… See the full description on the dataset page:
https://huggingface.co/datasets/cs552-the-expendables/patientagent-eval-results.