Do 3D Large Language Models Really Understand 3D Spatial Relationships?
π Project Page Β· π Paper Β· π» GitHub
Real-3DQA is a debiased 3D spatial QA benchmark with viewpoint rotation consistency evaluation. It addresses two key shortcomings of existing benchmarks:
Language Shortcut Filtering β Questions answerable through linguistic priors alone are removed by comparing 3D-LLMs against blind text-only counterparts.
Viewpoint Rotation Score (VRS) β Eachβ¦ See the full description on the dataset page:
https://huggingface.co/datasets/Oliver-Ma/Real-3DQA.