A labeled, multilingual (EN/RO) benchmark for scam/fraud message detection,
with a first-class hard-legitimate subset. ScamGuardBench is the evaluation
companion to the scam-guard detector. Its headline metric is false-positive
rate on legitimate messages — because every failed consumer scam filter dies of
false positives, not of missed scams.
A scam detector that flags real bank OTP messages gets disabled within a… See the full description on the dataset page:
https://huggingface.co/datasets/flowxai/scamguardbench.