CMI Pref Pseudo is the larger preference comparison dataset of cmi-pref for multimodal-prompted music generation research. With 56k generations from 23 music generation models and 165k pairwise comparisons, this dataset is designed to support research on music preference modeling and evaluation. 115k of the labels are labeled using qwen3-omni with our data-labeling pipeline. Prompts are compositional, with text, optinal lyrics and reference audio. The dataset is… See the full description on the dataset page:
https://huggingface.co/datasets/HaiwenXia/cmi-pref-pseudo.