BLUEX-v2 is a benchmark for evaluating Large Language Models on open-ended (discursive) questions from two of Brazil's most prestigious university entrance exams:
UNICAMP (Comvest) — University of Campinas
USP (Fuvest) — University of São Paulo
The dataset covers exam years 2022–2025 and focuses exclusively on the discursive (free-form answer) phase of these exams. Models are expected… See the full description on the dataset page:
https://huggingface.co/datasets/Tropic-AI/BLUEX-v2.