This is the resized version of the TACO dataset, with all allocentric videos and segmentation masks downscaled to a uniform 512x376 resolution (from native 4096x3000 / 2048x1500). Camera intrinsics are rescaled accordingly.
The original TACO allocentric videos are 4096x3000, making training impractical without on-the-fly resizing. This version… See the full description on the dataset page:
https://huggingface.co/datasets/mzhobro/taco_dataset_resized.