activity
20172022
most citedYOLOX: Exceeding YOLO Series in 2021

3k citations · 3.5k across the 20 of their papers we have counts for

collaborators
Showing cs.CVShow all

25 papers · 1 filter

cs.CV20222 cited

Generalizing Multiple Object Tracking to Unseen Domains by Introducing Natural Language Representation

En Yu, Songtao Liu, Zhuoling Li +4

Although existing multi-object tracking (MOT) algorithms have obtained competitive performance on various benchmarks, almost all of them train and validate models on the same domai…

cs.CV20223 cited

MatrixVT: Efficient Multi-Camera to BEV Transformation for 3D Perception

Hongyu Zhou, Zheng Ge, Zeming Li +1

This paper proposes an efficient multi-camera to Bird's-Eye-View (BEV) view transformation method for 3D perception, dubbed MatrixVT. Existing view transformers either suffer from…

cs.CV202219 cited

BEVStereo: Enhancing Depth Estimation in Multi-view 3D Object Detection with Dynamic Temporal Stereo

Yinhao Li, Han Bao, Zheng Ge +3

Bounded by the inherent ambiguity of depth perception, contemporary camera-based 3D object detection methods fall into the performance bottleneck. Intuitively, leveraging temporal…

cs.CV2022

Voxel Field Fusion for 3D Object Detection

Yanwei Li, Xiaojuan Qi, Yukang Chen +4

In this work, we present a conceptually simple yet effective framework for cross-modality 3D object detection, named voxel field fusion. The proposed approach aims to maintain cros…

cs.CV20221 cited

Real-time Object Detection for Streaming Perception

Jinrong Yang, Songtao Liu, Zeming Li +2

Autonomous driving requires the model to perceive the environment and (re)act within a low latency for safety. While past works ignore the inevitable changes in the environment aft…

cs.CV20213k cited

YOLOX: Exceeding YOLO Series in 2021

Zheng Ge, Songtao Liu, Feng Wang +2

In this report, we present some experienced improvements to YOLO series, forming a new high-performance detector -- YOLOX. We switch the YOLO detector to an anchor-free manner and…