Testing the ability of LLMs in finding more than one correct choices in a medical domains, where a penalty is added for incorrect ones, simulating real world evaluating scenarios
of medical students.
Large Language Models (LLMs) constitute a breakthrough state-of-the-art Artificial Intelligence (AI) technology which is rapidly evolving and promises to aid in medical diagnosis… See the full description on the dataset page:
https://huggingface.co/datasets/DimitriosPanagoulias/COGNET-MD.