13 citations · 17 across the 3 of their papers we have counts for
4 papers
iQuery: Instruments as Queries for Audio-Visual Sound Separation
Jiaben Chen, Renrui Zhang, Dongze Lian +3
Current audio-visual separation methods share a standard architecture design where an audio encoder-decoder network is fused with visual encoding features at the encoder bottleneck…
POS-BERT: Point Cloud One-Stage BERT Pre-Training
Kexue Fu, Peng Gao, ShaoLei Liu +3
Recently, the pre-training paradigm combining Transformer and masked language modeling has achieved tremendous success in NLP, images, and point clouds, such as BERT. However, dire…
Distillation with Contrast is All You Need for Self-Supervised Point Cloud Representation Learning
Kexue Fu, Peng Gao, Renrui Zhang +3
In this paper, we propose a simple and general framework for self-supervised point cloud representation learning. Human beings understand the 3D world by extracting two levels of i…
End-to-End Object Detection with Adaptive Clustering Transformer
Minghang Zheng, Peng Gao, Renrui Zhang +4
End-to-end Object Detection with Transformer (DETR)proposes to perform object detection with Transformer and achieve comparable performance with two-stage object detection like Fas…