This dataset is a relabeled subset derived from StarVector's
starvector/svg-stack dataset. It pairs SVG source code with generated natural
language captions describing the rendered visual appearance of each SVG.
The dataset is intended for supervised fine-tuning of text-to-SVG generation
models. A typical training format is to use Caption as the input prompt and
Svg as the target completion.
Each row contains:
Filename: unique SVG filename from the… See the full description on the dataset page:
https://huggingface.co/datasets/JackertheHacker/svg-stack-qwen-captioned-subset.