Views
No views yet
CC3M spoken captions (image–speech) — DualCodec pre-tokenized
Contents (train split)