To support community developers in avoiding the phenomenon of "catastrophic forgetting" when fine-tuning the DistilQwen2.5 model, we have open-sourced a portion of the dataset used for model training.
These datasets are designed to provide a solid foundation for model fine-tuning, helping to enhance the model's adaptability to new tasks while maintaining its performance on previous ones.
The released data covers various domains, including mathematics, coding, knowledge-based Q&A… See the full description on the dataset page:
https://huggingface.co/datasets/alibaba-pai/DistilQwen_100k.