JamendoMaxCaps is a large-scale dataset of over 362,000 instrumental tracks sourced from the Jamendo platform. It includes generated music captions and original metadata. Additionally, we introduce a retrieval system that utilizes both musical features and metadata to identify similar songs, which are then used to impute missing metadata via a local large language model (LLLM).
This dataset facilitates research in:
Music-language understanding
Music… See the full description on the dataset page:
https://huggingface.co/datasets/amaai-lab/JamendoMaxCaps.