This dataset is part of the SketchVLM project, introduced in the paper: SketchVLM: Vision language models can annotate images to explain thoughts and guide users.
Project Page | GitHub | Interactive Demo
SketchVLM is a training-free, model-agnostic framework that enables vision-language models (VLMs) to produce non-destructive, editable SVG overlays on input images to visually explain their reasoning.
The Physics Ball… See the full description on the dataset page:
https://huggingface.co/datasets/loganbolton/sketchvlm-physics-ball-drop.