activity
20202024
most citedDeep Multimodal Fusion by Channel Exchanging

118 citations · 129 across the 11 of their papers we have counts for

collaborators

8 papers

cs.CV2023

Root Pose Decomposition Towards Generic Non-rigid 3D Reconstruction with Monocular Videos

Yikai Wang, Yinpeng Dong, Fuchun Sun +1

This work focuses on the 3D reconstruction of non-rigid objects based on monocular RGB video sequences. Concretely, we aim at building high-fidelity models for generic object categ…

cs.CV20231 cited

Human-imperceptible, Machine-recognizable Images

Fusheng Hao, Fengxiang He, Yikai Wang +4

Massive human-related data is collected to train neural networks for computer vision tasks. A major conflict is exposed relating to software engineers between better developing AI…

cs.CV2023

Compacting Binary Neural Networks by Sparse Kernel Selection

Yikai Wang, Wenbing Huang, Yinpeng Dong +2

Binary Neural Network (BNN) represents convolution weights with 1-bit values, which enhances the efficiency of storage and computation. This paper is motivated by a previously reve…

cs.CV20231 cited

Benchmarking Robustness of 3D Object Detection to Common Corruptions in Autonomous Driving

Yinpeng Dong, Caixin Kang, Jinlai Zhang +6

3D object detection is an important task in autonomous driving to perceive the surroundings. Despite the excellent performance, the existing 3D detectors lack the robustness to rea…

cs.CV20222 cited

Bridged Transformer for Vision and Point Cloud 3D Object Detection

Yikai Wang, TengQi Ye, Lele Cao +4

3D object detection is a crucial research topic in computer vision, which usually uses 3D point clouds as input in conventional setups. Recently, there is a trend of leveraging mul…

cs.SD20227 cited

Sound Adversarial Audio-Visual Navigation

Yinfeng Yu, Wenbing Huang, Fuchun Sun +3

Audio-visual navigation task requires an agent to find a sound source in a realistic, unmapped 3D environment by utilizing egocentric audio-visual observations. Existing audio-visu…