Web Content Classification Dataset with HTML and Screenshots
Dataset
Description: This dataset contains 1,000 carefully curated examples of web content designed for phishing detection classification tasks. Each example consists of a website URL along with its complete HTML content, a visual screenshot, and a manually verified classification label.
Important Note: This entire dataset was meticulously created through manual processes. Each website was individually visited… See the full description on the dataset page: https://huggingface.co/datasets/mafiabasbush/Dataset_Phising.