Views
No views yet
Shengli Zhou, Minghang Zheng, Feng Zheng, and Yang Liu. 2026. Scalable Object Relation Encoding for Better 3D Spatial Reasoning in Large Language Models. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2026.
| ScanRefer | Multi3DRefer | SQA3D | |||
|---|---|---|---|---|---|
| Model | Acc.@0.25 | Acc.@0.5 | F1@0.25 | F1@0.5 | EM@1 |
| Chat-Scene + QuatRoPE | 57.8 | 52.2 | 59.5 | 54.8 | 54.7 |
| 3DGraphLLM + QuatRoPE | 58.2 | 52.5 | 60.6 | 56.0 | 55.2 |
1@misc{zhou2026scalableobjectrelationencoding,
2 title={Scalable Object Relation Encoding for Better 3D Spatial Reasoning in Large Language Models},
3 author={Shengli Zhou and Minghang Zheng and Feng Zheng and Yang Liu},
4 year={2026},
5 eprint={2603.24721},
6 archivePrefix={arXiv},
7 primaryClass={cs.CV},
8 url={https://arxiv.org/abs/2603.24721},
9}