Dataset type:
LLaVA Visual Instruct 150K is a set of GPT-generated multimodal instruction-following data.
It is constructed for visual instruction tuning and for building large multimodal towards GPT-4 vision/language capability.
Dataset… See the full description on the dataset page: https://huggingface.co/datasets/gtz1/LLaVA-Instruct-150K.