11 citations · 15 across the 3 of their papers we have counts for
3 papers
cs.RO2023★ 3 cited
HuBo-VLM: Unified Vision-Language Model designed for HUman roBOt interaction tasks
Zichao Dong, Weikun Zhang, Xufeng Huang +3
Human robot interaction is an exciting task, which aimed to guide robots following instructions from human. Since huge gap lies between human natural language and machine codes, en…
cs.CV2023★ 11 cited
FusionAD: Multi-modality Fusion for Prediction and Planning Tasks of Autonomous Driving
Tengju Ye, Wei Jing, Chunyong Hu +11
Building a multi-modality multi-task neural network toward accurate and robust performance is a de-facto standard in perception task of autonomous driving. However, leveraging such…
cs.CV2023★ 1 cited
OG: Equip vision occupancy with instance segmentation and visual grounding
Zichao Dong, Hang Ji, Weikun Zhang +2
Occupancy prediction tasks focus on the inference of both geometry and semantic labels for each voxel, which is an important perception mission. However, it is still a semantic seg…