most citedDetCLIP: Dictionary-Enriched Visual-Concept Paralleled Pre-training for Open-world Detection

64 citations · 141 across the 16 of their papers we have counts for

collaborators

17 papers

cs.CV202264 cited

DetCLIP: Dictionary-Enriched Visual-Concept Paralleled Pre-training for Open-world Detection

Lewei Yao, Jianhua Han, Youpeng Wen +6

Open-world object detection, as a more general and challenging goal, aims to recognize and localize objects described by arbitrary category names. The recent work GLIP formulates t…

cs.CV202216 cited

Effective Adaptation in Multi-Task Co-Training for Unified Autonomous Driving

Xiwen Liang, Yangxin Wu, Jianhua Han +3

Aiming towards a holistic understanding of multiple downstream tasks simultaneously, there is a need for extracting features with better transferability. Though many latest self-su…

cs.CV2022

Task-Customized Self-Supervised Pre-training with Scalable Dynamic Routing

Zhili Liu, Jianhua Han, Lanqing Hong +4

Self-supervised learning (SSL), especially contrastive methods, has raised attraction recently as it learns effective transferable representations without semantic annotations. A c…

cs.CV20222 cited

Continual Object Detection via Prototypical Task Correlation Guided Gating Mechanism

Binbin Yang, Xinchi Deng, Han Shi +6

Continual learning is a challenging real-world problem for constructing a mature AI system when data are provided in a streaming fashion. Despite recent progress in continual class…

cs.CV20224 cited

ManiTrans: Entity-Level Text-Guided Image Manipulation via Token-wise Semantic Alignment and Generation

Jianan Wang, Guansong Lu, Hang Xu +3

Existing text-guided image manipulation methods aim to modify the appearance of the image or to edit a few objects in a virtual or simple scenario, which is far from practical appl…

cs.CV20222 cited

Point2Seq: Detecting 3D Objects as Sequences

Yujing Xue, Jiageng Mao, Minzhe Niu +5

We present a simple and effective framework, named Point2Seq, for 3D object detection from point clouds. In contrast to previous methods that normally {predict attributes of 3D obj…