Source: the authors' Hugging Face backup yhcao/V3Det_Backup (annotation JSONs + 21 image zips, 53.5 GB), linked as the HuggingFace download channel from github.com/V3Det/V3Det.
Converted by the finedet project into a unified, AutoTrain-compatible layout:
image / width / height / objects{bbox, category} with COCO-format
[x, y, w, h] boxes in absolute pixels. Boxes are clipped to the image and
empty boxes… See the full description on the dataset page:
https://huggingface.co/datasets/finedet/v3det.