MultiFin – a publicly available financial dataset consisting of real-world article headlines covering 15 languages across different writing systems and language families.
The dataset consists of hierarchical label structure providing two classification tasks: multi-label and multi-class.
The MULTIFIN dataset is a multilingual corpus, consisting of real-world article headlines covering 15
languages. The corpus is annotated using hierarchical… See the full description on the dataset page:
https://huggingface.co/datasets/awinml/MultiFin.