Wikipedia Personas is a dataset constructed from paragraphs sampled from the agentlans/wikipedia-paragraphs-complete dataset
using the sample_k10000 and sample_k20000 configurations.
Each paragraph is paired with a persona crafted as a plausible expert, enthusiast, or stakeholder related to the content of the corresponding Wikipedia text.
Personas were initially seeded with 20 handcrafted examples following the style of proj-persona/PersonaHub
and then expanded… See the full description on the dataset page:
https://huggingface.co/datasets/agentlans/wikipedia-personas.