Dataset used for paper -> "Rethinking Dataset Compression: Shifting Focus From Labels to Images"
Dataset created according to the paper Identifying Mislabeled Data using the Area Under the Margin Ranking.
from datasets import load_dataset
dataset = load_dataset("he-yang/2025-rethinkdc-imagenet-aum-ipc-10")
For more information, please refer to the Rethinking-Dataset-Compression