activity
20192023
most citedSOLQ: Segmenting Objects by Learning Queries

56 citations · 112 across the 9 of their papers we have counts for

collaborators

12 papers

cs.CV20231 cited

Panacea: Panoramic and Controllable Video Generation for Autonomous Driving

Yuqing Wen, Yucheng Zhao, Yingfei Liu +7

The field of autonomous driving increasingly demands high-quality annotated training data. In this paper, we propose Panacea, an innovative approach to generate panoramic and contr…

cs.CV20237 cited

ADriver-I: A General World Model for Autonomous Driving

Fan Jia, Weixin Mao, Yingfei Liu +5

Typically, autonomous driving adopts a modular design, which divides the full stack into perception, prediction, planning and control parts. Though interpretable, such modular desi…

cs.CV20231 cited

VLM-Eval: A General Evaluation on Video Large Language Models

Shuailin Li, Yuang Zhang, Yucheng Zhao +4

Despite the rapid development of video Large Language Models (LLMs), a comprehensive evaluation is still absent. In this paper, we introduce a unified evaluation that encompasses m…

cs.CV2022

Towards 3D Object Detection with 2D Supervision

Jinrong Yang, Tiancai Wang, Zheng Ge +3

The great progress of 3D object detectors relies on large-scale data and 3D annotations. The annotation cost for 3D bounding boxes is extremely expensive while the 2D ones are easi…

cs.CV2022

The 1st-place Solution for ECCV 2022 Multiple People Tracking in Group Dance Challenge

Yuang Zhang, Tiancai Wang, Weiyao Lin +1

We present our 1st place solution to the Group Dance Multiple People Tracking Challenge. Based on MOTR: End-to-End Multiple-Object Tracking with Transformer, we explore: 1) detect…

cs.CV20224 cited

Tree Energy Loss: Towards Sparsely Annotated Semantic Segmentation

Zhiyuan Liang, Tiancai Wang, Xiangyu Zhang +2

Sparsely annotated semantic segmentation (SASS) aims to train a segmentation network with coarse-grained (i.e., point-, scribble-, and block-wise) supervisions, where only a small…