Want to fine-tune this dataset on LLaMA-Factory? Check this repository for preprocessing: llm-merging datasets
I automatically converted the dataset into the default format that can be previewed on huggingface.
In this work, we present the first free-form multiple-choice OpenQA dataset for solving medical problems, MedQA,
collected from the professional medical board exams. It covers three languages: English, simplified Chinese, and
traditional Chinese, and… See the full description on the dataset page:
https://huggingface.co/datasets/fzkuji/MedQA.