CyberSec-Bench is a bilingual (English/French) benchmark dataset designed to evaluate the cybersecurity knowledge of Large Language Models (LLMs) and AI systems. The dataset contains 200 expert-crafted questions spanning five critical domains of cybersecurity, with detailed reference answers for each question.
This benchmark tests real-world cybersecurity knowledge at professional… See the full description on the dataset page: https://huggingface.co/datasets/AYI-NEDJIMI/CyberSec-Bench.