2 papers
cs.AI2025
DeepThink3D: Enhancing Large Language Models with Programmatic Reasoning in Complex 3D Situated Reasoning Tasks
Jiayi Song, Rui Wan, Lipeng Ma +4
This work enhances the ability of large language models (LLMs) to perform complex reasoning in 3D scenes. Recent work has addressed the 3D situated reasoning task by invoking tool…
cs.CV2023
PTA-Det: Point Transformer Associating Point cloud and Image for 3D Object Detection
Rui Wan, Tianyun Zhao, Wei Zhao
In autonomous driving, 3D object detection based on multi-modal data has become an indispensable approach when facing complex environments around the vehicle. During multi-modal de…