A 10,000-sample, diversity-focused subset of NVIDIA’s Nemotron-Personas-USA synthetic persona dataset, curated to provide a compact but representative slice of U.S.-aligned synthetic personas for research and model development.
Name: Diverse Nemotron-Personas-USA 10K Subset
Source dataset: nvidia/Nemotron-Personas-USA
Records: 10,000 personas
Modality: Text
Primary language: English
License:… See the full description on the dataset page:
https://huggingface.co/datasets/anlee-0618/Nemotron-Personas-USA-10K.