This repository provides a processed version of the DISC-Law-SFT dataset for easier usage in instruction tuning and aligned language model training. The dataset has been converted into the Alpaca format, which is commonly used for supervised fine-tuning of language models on instruction-following tasks.
The original DISC-Law-SFT dataset was proposed for developing intelligent legal service systems with… See the full description on the dataset page:
https://huggingface.co/datasets/alfonsusrr/DISC-Law-SFT-Alpaca.