6 citations · 11 across the 7 of their papers we have counts for
7 papers
Data-efficient Event Camera Pre-training via Disentangled Masked Modeling
Zhenpeng Huang, Chao Li, Hao Chen +3
In this paper, we present a new data-efficient voxel-based self-supervised learning method for event cameras. Our pre-training overcomes the limitations of previous methods, which…
Tracking with Human-Intent Reasoning
Jiawen Zhu, Zhi-Qi Cheng, Jun-Yan He +5
Advances in perception modeling have significantly improved the performance of object tracking. However, the current methods for specifying the target object in the initial frame a…
FMViT: A multiple-frequency mixing Vision Transformer
Wei Tan, Yifeng Geng, Xuansong Xie
The transformer model has gained widespread adoption in computer vision tasks in recent times. However, due to the quadratic time and memory complexity of self-attention, which is…
PoSynDA: Multi-Hypothesis Pose Synthesis Domain Adaptation for Robust 3D Human Pose Estimation
Hanbing Liu, Jun-Yan He, Zhi-Qi Cheng +8
Existing 3D human pose estimators face challenges in adapting to new datasets due to the lack of 2D-3D pose pairs in training sets. To overcome this issue, we propose \textit{Multi…
FastInst: A Simple Query-Based Model for Real-Time Instance Segmentation
Junjie He, Pengyu Li, Yifeng Geng +1
Recent attention in instance segmentation has focused on query-based models. Despite being non-maximum suppression (NMS)-free and end-to-end, the superiority of these models on hig…
DAMO-StreamNet: Optimizing Streaming Perception in Autonomous Driving
Jun-Yan He, Zhi-Qi Cheng, Chenyang Li +5
Real-time perception, or streaming perception, is a crucial aspect of autonomous driving that has yet to be thoroughly explored in existing research. To address this gap, we presen…