This data is delivered under GNU Public License and you may use it freely.
To use it, download the ZIP file containing the XML corpus.
The following paper must be cited when using this corpus:
Taulé, M., M.A. Martí, M. Recasens (2008) 'Ancora: Multilevel Annotated Corpora for Catalan and Spanish',
Proceedings of 6th International Conference on Language Resources and Evaluation. Marrakesh (Morocco).
In addition, the following paper must be cited if coreference information (attributes entity… See the full description on the dataset page:
https://huggingface.co/datasets/CLiC-UB/AnCora-CA.