Evidence-grounded Speaker Card corpus for in-the-wild speaker verification.
SpeakerCard-1M is a speaker-centric resource built on VoxCeleb1/2 under a tool-first, LLM-last pipeline: ten acoustic probes extract field-level evidence, a schema separates relatively stable traits (gender, age band, accent, pitch band, timbre, language) from utterance-level states (emotion, channel, environment, speaking rate), and a constrained LLM verbalizes the… See the full description on the dataset page:
https://huggingface.co/datasets/JYP2024/SpeakerCard-1M.