⚠️ Warning: This dataset contains harmful and toxic language. It is intended for academic purposes only.
[ArXiv Paper] [Github] [Hugging Face Leaderboard 🤗]
The ThaiSafetyBench dataset comprises 1,889 malicious Thai-language prompts across various categories. In addition to translated malicious prompts, it includes prompts tailored to Thai culture, offering deeper insights into culturally specific attacks.
Note: The Monarchy type of harm has been removed from the… See the full description on the dataset page:
https://huggingface.co/datasets/typhoon-ai/ThaiSafetyBench.