Per-pair predictions for every model evaluated in An Analysis of the Performance of Large Language
Models in Spanish NLI Datasets with Causal Relationships (IBERAMIA 2026, to appear). Code in
Pacolas/NLI-via-LLM; part of the
ESNLIR-LLM
collection.
These are the raw outputs behind the paper's tables, so results can be re-scored, sliced by genre or
domain, or compared pair by pair without re-running any model.
Files… See the full description on the dataset page: https://huggingface.co/datasets/Flaglab/esnlir-llm-predictions.