Updates:
13th April : Caption files for dataset in Part_Aa are available here. Download separately and place inside Part_Aa directory
Usage:
The dataset is split into 10 .tar chunks, dowload and reassemble them using the steps below:
pip install huggingface_hub
from huggingface_hub import snapshot_download
Specify the repository ID and the local directory path