This dataset is being used as the training set for the S23DR Challenge.
This is a living dataset. Today, we provide 4316 samples for training, and 175 for validation and hold back an additional 1072 for computing the private and public leaderboards. Additional, we intend to continue releasing training data throughout the challenge and beyond. The data take the following form:
Features({
"order_id": Value(dtype="string"),
# inputs
"K":… See the full description on the dataset page:
https://huggingface.co/datasets/usm3d/hoho-train-set.