activity
20192024
most citedCrowdCLIP: Unsupervised Crowd Counting via Vision-Language Model

4 citations · 9 across the 9 of their papers we have counts for

collaborators

18 papers

cs.CV2024

UniDet: Unified and Universal Framework for Prompt-Guided Multi-dataset 3D Detection

Yubin Wang, Zhikang Zou, Xiaoqing Ye +3

We present UniDet, a brand new framework for unified and universal multi-dataset training on 3D detection, enabling robust performance across diverse domains and generalization…

cs.CV2024

Exploring the Causality of End-to-End Autonomous Driving

Jiankun Li, Hao Li, Jiangjiang Liu +6

Deep learning-based models are widely deployed in autonomous driving areas, especially the increasingly noticed end-to-end solutions. However, the black-box property of these model…

cs.CV2024

PointMamba: A Simple State Space Model for Point Cloud Analysis

Dingkang Liang, Xin Zhou, Wei Xu +5

Transformers have become one of the foundational architectures in point cloud analysis tasks due to their excellent global modeling ability. However, the attention mechanism has qu…

cs.CV2023

Diffusion-based 3D Object Detection with Random Boxes

Xin Zhou, Jinghua Hou, Tingting Yao +6

3D object detection is an essential task for achieving autonomous driving. Existing anchor-based detection methods rely on empirical heuristics setting of anchors, which makes the…

cs.CV2023

CityTrack: Improving City-Scale Multi-Camera Multi-Target Tracking by Location-Aware Tracking and Box-Grained Matching

Jincheng Lu, Xipeng Yang, Jin Ye +4

Multi-Camera Multi-Target Tracking (MCMT) is a computer vision technique that involves tracking multiple targets simultaneously across multiple cameras. MCMT in urban traffic visua…

cs.CV2023

Understanding Depth Map Progressively: Adaptive Distance Interval Separation for Monocular 3d Object Detection

Xianhui Cheng, Shoumeng Qiu, Zhikang Zou +2

Monocular 3D object detection aims to locate objects in different scenes with just a single image. Due to the absence of depth information, several monocular 3D detection technique…