Seneca-CyBench: A comprehensive benchmark system designed to evaluate Large Language Models (LLMs) on cybersecurity domain knowledge. Features GPT-4o-based automated scoring for objective assessment of model capabilities across security topics.
620 questions (310 MCQ + 310 SAQ) covering all major cybersecurity domains including GRC, Security Architecture, Cloud Security, IAM, and more.