3k citations · 3.7k across the 18 of their papers we have counts for
23 papers
Align-DETR: Enhancing End-to-end Object Detection with Aligned Loss
Zhi Cai, Songtao Liu, Guodong Wang +3
DETR has set up a simple end-to-end pipeline for object detection by formulating this task as a set prediction problem, showing promising potential. Despite its notable advancement…
Dynamic Grained Encoder for Vision Transformers
Lin Song, Songyang Zhang, Songtao Liu +5
Transformers, the de-facto standard for language modeling, have been recently applied for vision tasks. This paper introduces sparse queries for vision transformers to exploit the…
Generalizing Multiple Object Tracking to Unseen Domains by Introducing Natural Language Representation
En Yu, Songtao Liu, Zhuoling Li +4
Although existing multi-object tracking (MOT) algorithms have obtained competitive performance on various benchmarks, almost all of them train and validate models on the same domai…
StreamYOLO: Real-time Object Detection for Streaming Perception
Jinrong Yang, Songtao Liu, Zeming Li +2
The perceptive models of autonomous driving require fast inference within a low latency for safety. While existing works ignore the inevitable environmental changes after processin…
Dense Teacher: Dense Pseudo-Labels for Semi-supervised Object Detection
Hongyu Zhou, Zheng Ge, Songtao Liu +4
To date, the most powerful semi-supervised object detectors (SS-OD) are based on pseudo-boxes, which need a sequence of post-processing with fine-tuned hyper-parameters. In this wo…
DBQ-SSD: Dynamic Ball Query for Efficient 3D Object Detection
Jinrong Yang, Lin Song, Songtao Liu +6
Many point-based 3D detectors adopt point-feature sampling strategies to drop some points for efficient inference. These strategies are typically based on fixed and handcrafted rul…