CaReSound is a benchmark dataset designed for open-ended diagnostic reasoning using cardiac and respiratory auscultation audio. It includes annotated medical sound recordings enriched with metadata and automatically generated question-answer (QA) pairs, enabling research into audio-language modeling for healthcare.
Multimodal: Each sample pairs an auscultation audio clip with metadata and diagnostic question-answer pairs.
Diverse: Built from… See the full description on the dataset page:
https://huggingface.co/datasets/tsnngw/CaReSound.