Are Vision-Language Models Safe in the Wild? A Meme-Based Benchmark Study
MemeSafetyBench: A Benchmark for Assessing VLM Safety with Real-World Memes
🎉 09/03/2025: We've also released MemeSafetyBench-Mini, a lightweight benchmark consisting of 390 samples across 13 categories (30 samples per category).
🎉 09/03/2025: Our code is now available on GitHub!
🎉 08/21/2025: MemeSafetyBench is accepted at EMNLP 2025!… See the full description on the dataset page:
https://huggingface.co/datasets/oneonlee/Meme-Safety-Bench.