activity
20142024
most citedTextBoxes: A Fast Text Detector with a Single Deep Neural Network

445 citations · 660 across the 34 of their papers we have counts for

collaborators

26 papers

cs.CV20242 cited

Not All Voxels Are Equal: Hardness-Aware Semantic Scene Completion with Self-Distillation

Song Wang, Jiawei Yu, Wentong Li +4

Semantic scene completion, also known as semantic occupancy prediction, can provide dense geometric and semantic information for autonomous vehicles, which attracts the increasing…

cs.CV2024

MIM4D: Masked Modeling with Multi-View Video for Autonomous Driving Representation Learning

Jialv Zou, Bencheng Liao, Qian Zhang +2

Learning robust and scalable visual representations from massive multi-view video data remains a challenge in computer vision and autonomous driving. Existing pre-training methods…

cs.CV202428 cited

YOLO-World: Real-Time Open-Vocabulary Object Detection

Tianheng Cheng, Lin Song, Yixiao Ge +3

The You Only Look Once (YOLO) series of detectors have established themselves as efficient and practical tools. However, their reliance on predefined and trained object categories…

cs.CV2023

Circuit as Set of Points

Jialv Zou, Xinggang Wang, Jiahao Guo +3

As the size of circuit designs continues to grow rapidly, artificial intelligence technologies are being extensively used in Electronic Design Automation (EDA) to assist with circu…

cs.CV20233 cited

Label-efficient Segmentation via Affinity Propagation

Wentong Li, Yuqian Yuan, Song Wang +5

Weakly-supervised segmentation with label-efficient sparse annotations has attracted increasing research attention to reduce the cost of laborious pixel-wise labeling process, whil…

cs.CV2023

Condition-Adaptive Graph Convolution Learning for Skeleton-Based Gait Recognition

Xiaohu Huang, Xinggang Wang, Zhidianqiu Jin +4

Graph convolutional networks have been widely applied in skeleton-based gait recognition. A key challenge in this task is to distinguish the individual walking styles of different…