This dataset contains approximately 164,000 YouTube thumbnails paired with their corresponding video titles.
The dataset was constructed by collecting public YouTube channel feeds, extracting video metadata, filtering and deduplicating entries, and downloading thumbnail images at scale.
The goal of this dataset is to support research and experimentation in: