This dataset is designed for instruction tuning vision-language models (VLMs) on 3D Computer-Aided Design (CAD) comprehension and interactive engineering reasoning. It scales up the Text2CAD dataset dataset by transforming static CAD assets and multi-level design prompts into a multimodal, conversational format.
To bridge the gap between static 3D CAD data and conversational AI, we extended the original dataset by:
Rendering multi-view images for each CAD asset… See the full description on the dataset page:
https://huggingface.co/datasets/trislee02/text2cad_multiview.