MedChatQA dataset aims to be a benchmark for testing LLMs for accurate QA on real-world Medical Information and Medical Communication topics.
There are several professionals in the medical field who communicate with patients, and with other professionals in their field.
These communications are expected to be 100% factual and free of errors.
The MedChatQA Dataset aims to help anyone building GenAI products in the medical vertical to test and… See the full description on the dataset page:
https://huggingface.co/datasets/ngram/medchat-qa.