Meetween YouTube Meeting Video Dataset
Dataset Description
The Meetween YouTube Meeting Dataset is a comprehensive corpus comprising processed transcripts and metadata from 122,407 public YouTube videos, totaling over 80,000 hours of audiovisual content. Part of the Mumospee v2 release, this dataset is specifically designed to support AI research in meeting analysis, speech processing, and multimodal technologies.
The content focuses on virtual meetings, webinars… See the full description on the dataset page: https://huggingface.co/datasets/meetween/meetween_youtube_meeting.