This dataset is designed for fine-tuning vision-language models, specifically PaliGemma, to perform medical visual question answering (VQA). It contains real-world medical images paired with patient-style questions and doctor-style answers.
✅ Multimodal: Images + Text
✅ Doctor-style professional answers
✅ Focused on common medical conditions
✅ Suitable for LoRA fine-tuning and… See the full description on the dataset page:
https://huggingface.co/datasets/SiyunHE/medical-pilagemma-lora.