Bongard Problems (BPs) provide a challenging testbed for abstract visual reasoning (AVR), requiring models to identify visual concepts from just a few examples and describe them in natural language.
Early BP benchmarks featured synthetic black-and-white drawings, which might not fully capture the complexity of real-world scenes.
Subsequent BP datasets employed real-world images, albeit
%with real-world
the represented concepts are identifiable from high-level image… See the full description on the dataset page:
https://huggingface.co/datasets/spawlonka/bongard-rwr-plus.