This is the repository for PLOD Dataset subset being used for CW in NLP module 2023-2024 at University of Surrey.
Dataset Summary
This PLOD Dataset is an English-language dataset of abbreviations and their long-forms tagged in text. The dataset has been collected for research from the PLOS journals indexing of abbreviations and long-forms in the text. This dataset was created to support the Natural Language Processing task of… See the full description on the dataset page: https://huggingface.co/datasets/surrey-nlp/PLOD-CW.