56 citations · 112 across the 9 of their papers we have counts for
12 papers
Panacea: Panoramic and Controllable Video Generation for Autonomous Driving
Yuqing Wen, Yucheng Zhao, Yingfei Liu +7
The field of autonomous driving increasingly demands high-quality annotated training data. In this paper, we propose Panacea, an innovative approach to generate panoramic and contr…
ADriver-I: A General World Model for Autonomous Driving
Fan Jia, Weixin Mao, Yingfei Liu +5
Typically, autonomous driving adopts a modular design, which divides the full stack into perception, prediction, planning and control parts. Though interpretable, such modular desi…
VLM-Eval: A General Evaluation on Video Large Language Models
Shuailin Li, Yuang Zhang, Yucheng Zhao +4
Despite the rapid development of video Large Language Models (LLMs), a comprehensive evaluation is still absent. In this paper, we introduce a unified evaluation that encompasses m…
Towards 3D Object Detection with 2D Supervision
Jinrong Yang, Tiancai Wang, Zheng Ge +3
The great progress of 3D object detectors relies on large-scale data and 3D annotations. The annotation cost for 3D bounding boxes is extremely expensive while the 2D ones are easi…
The 1st-place Solution for ECCV 2022 Multiple People Tracking in Group Dance Challenge
Yuang Zhang, Tiancai Wang, Weiyao Lin +1
We present our 1st place solution to the Group Dance Multiple People Tracking Challenge. Based on MOTR: End-to-End Multiple-Object Tracking with Transformer, we explore: 1) detect…
Tree Energy Loss: Towards Sparsely Annotated Semantic Segmentation
Zhiyuan Liang, Tiancai Wang, Xiangyu Zhang +2
Sparsely annotated semantic segmentation (SASS) aims to train a segmentation network with coarse-grained (i.e., point-, scribble-, and block-wise) supervisions, where only a small…