This repository contains a dataset of Traditional Chinese Medicine (TCM) for fine-tuning large language models.
The dataset contains 7,096 Chinese sentences related to TCM. The sentences are collected from various sources on the Internet, including medical websites, TCM forums, and TCM books. The dataset is generated or judged by various LLMs, including… See the full description on the dataset page:
https://huggingface.co/datasets/Monor/hwtcm-sft-v1.