3 citations · 6 across the 4 of their papers we have counts for
4 papers
RepVF: A Unified Vector Fields Representation for Multi-task 3D Perception
Chunliang Li, Wencheng Han, Junbo Yin +2
Concurrent processing of multiple autonomous driving 3D perception tasks within the same spatiotemporal scene poses a significant challenge, in particular due to the computational…
InteractiveVideo: User-Centric Controllable Video Generation with Synergistic Multimodal Instructions
Yiyuan Zhang, Yuhao Kang, Zhixin Zhang +3
We introduce , a user-centric framework for video generation. Different from traditional generative approaches that operate based on user-provided images…
Generalized Few-Shot 3D Object Detection of LiDAR Point Cloud for Autonomous Driving
Jiawei Liu, Xingping Dong, Sanyuan Zhao +1
Recent years have witnessed huge successes in 3D object detection to recognize common objects for autonomous driving (e.g., vehicles and pedestrians). However, most methods rely he…
Bilateral Cross-Modality Graph Matching Attention for Feature Fusion in Visual Question Answering
JianJian Cao, Xiameng Qin, Sanyuan Zhao +1
Answering semantically-complicated questions according to an image is challenging in Visual Question Answering (VQA) task. Although the image can be well represented by deep learni…