train: Alpaca-format data for instruction tuning from Openhermes_10w
validation: multiple-choice reasoning data for evaluation, which includes not only the ground-truth answers but also model-generated predictions
A high-quality instruction–response dataset in Alpaca format, designed for training and fine-tuning large language models on instruction-following… See the full description on the dataset page:
https://huggingface.co/datasets/OpenDCAI/DataFlex-selector-openhermes-10w.