JA-VLM-Bench-In-the-Wild is Japanese version of LLaVA-Bench-In-the-Wild.
We carefully collected a diverse set of 42 images with 50 questions in total. (For LLaVA-Bench-In-the-Wild, 24 images with 60 questions)
The images contain Japanese culture and objects in Japan. The Japanese questions and answers were generated with assistance from GPT-4V (gpt-4-vision-preview), OpenAI’s large-scale language-generation model and removed… See the full description on the dataset page: https://huggingface.co/datasets/SakanaAI/JA-VLM-Bench-In-the-Wild.