This dataset is released as part of the paper "Sub-Billion, Super-Frontier: Fine-Tuned Small Language Models Rival Zero-Shot Frontier LLMs on General and Literary Relation Extraction" (Christou & Tsoumakas, 2026) arXiv:2606.22606.
The dataset is a processed version of the DFKI-SLT/conll04 dataset, tailored for relation extraction tasks. The original dataset, based on the CoNLL-2004 shared task… See the full description on the dataset page: https://huggingface.co/datasets/Despina/conll04.