SVGX-SFT-1M is a large-scale multimodal dataset designed for training and evaluating models on SVG-based
instruction-following and dialogue tasks.
It contains over 1 million samples covering:
Text-to-SVG Generation: Given a natural language prompt, generate corresponding SVG code.
SVG-to-Text Description: Given SVG code, produce a natural language description.
Image-to-SVG Reasoning (with ShareGPT-style conversations): Given a raster image, generate SVG code… See the full description on the dataset page:
https://huggingface.co/datasets/xingxm/SVGX-SFT-1M.