This dataset consists of synthetically generated question-answer pairs designed to simulate the process of answering high-level research questions about biomedical papers. It was created using OpenAI's GPT-4o model and is tailored for fine-tuning or evaluating models on tasks such as biomedical reading comprehension, information extraction, and reasoning.
Each data sample is a JSON… See the full description on the dataset page:
https://huggingface.co/datasets/AbrehamT/classified_papers.