i1: A Simple and Fully Open Recipe for Strong Text-to-Image Models
Boya Zeng, Tianze Luo, Shu Pu, Jucheng Shen, Taiming Lu, Gabriel Sarch, Zhuang Liu
Princeton University
[arXiv][code][model][project page]
To prepare the dataset for training, we store the image-caption pairs as TFRecords.
This HuggingFace dataset contains the TFRecords corresponding to the pexels dataset at 512×512 resolution. Concretely, we only retain raw images with a shorter edge of at least… See the full description on the dataset page:
https://huggingface.co/datasets/i1-datasets/i1-pexels-512-resolution-1m-tfrecord.