A benchmark of 1,300 synthetic prompts with 4,014 ground-truth
annotations spanning four workload classes, designed to evaluate
privacy-preserving techniques for outbound LLM requests.
Released alongside the paper:
LLM-Redactor: An Empirical Evaluation of Eight Techniques for
Privacy-Preserving LLM Requests
Justice Owusu Agyemang, Jerry John Kponyo, Elliot Amponsah,
Godfred Manu Addo Boakye, Kwame Opuni-Boachie Obour Agyekum
arXiv:2604.12064… See the full description on the dataset page:
https://huggingface.co/datasets/jayluxferro/llm-redactor-leak-benchmark.