This dataset has been created with distilabel.
The pipeline script was uploaded to easily reproduce the dataset:
text_classification.py.
It can be run directly using the CLI:
distilabel pipeline run --script "
https://huggingface.co/datasets/dvilasuero/synth-text-classification/raw/main/text_classification.py"
This dataset contains a pipeline.yaml which can be used to reproduce the pipeline that… See the full description on the dataset page:
https://huggingface.co/datasets/dvilasuero/synth-text-classification.