Paper | Project Page | Code
A subset from the 3D-Alpaca dataset of ShapeLLM-Omni: a native multimodal LLM for 3D generation and understanding
Junliang Ye*, Zhengyi Wang*, Ruowen Zhao*, Shenghao Xie, Jun Zhu
Recently, the powerful text-to-image capabilities of GPT-4o have led to growing appreciation for native multimodal large language models. However, its multimodal capabilities remain confined to images… See the full description on the dataset page:
https://huggingface.co/datasets/yejunliang23/3D-Alpaca.