The model has been fine-tuned using the Clinical Case and Question + automatically obtained RAG using
the MedCorp and MedRAG method
with 32 snippets. The model generates as output a prediction of the correct answer to the multiple choice exam and has been evaluated on 4 languages: English, French, Italian and Spanish.
For details about fine-tuning and evaluation please check the paper and the repository for usage.
1@misc{alonso2024medexpqa,
2 title={MedExpQA: Multilingual Benchmarking of Large Language Models for Medical Question Answering},
3 author={Iñigo Alonso and Maite Oronoz and Rodrigo Agerri},
4 year={2024},
5 eprint={2404.05590},
6 archivePrefix={arXiv},
7 primaryClass={cs.CL}
8}