This dataset, derived from VG150, provides image-text pairs for scene graph generation. Each example includes an image, an "open" prompt, a "close" prompt, a list of objects, and their relationships. It's designed to be used for training and evaluating models that generate scene graphs from images and textual prompts.
This dataset is used in the paper R1-SGG: Compile Scene Graphs with Reinforcement Learning.
The dataset is structured as follows:
image_id: Unique identifier for the image.… See the full description on the dataset page:
https://huggingface.co/datasets/JosephZ/vg150_val_sgg_prompt.