A benchmark dataset for evaluating Incomplete Prompt Jailbreak (IPJ) behavior in LLMs.
This dataset provides harmful-question seeds, paraphrase metadata, and fixed trigger templates for reproducible completion-style and chat-template-style safety evaluation.
chat_template_benchmark rows
1,890… See the full description on the dataset page:
https://huggingface.co/datasets/leo-bjpark/incomplete-prompt-jailbreak.