5 citations · 5 across the 2 of their papers we have counts for
2 papers
cs.CV2023
Deep Neural Networks in Video Human Action Recognition: A Review
Zihan Wang, Yang Yang, Zhi Liu +1
Currently, video behavior recognition is one of the most foundational tasks of computer vision. The 2D neural networks of deep learning are built for recognizing pixel-level inform…
cs.CV2023★ 5 cited
PiMAE: Point Cloud and Image Interactive Masked Autoencoders for 3D Object Detection
Anthony Chen, Kevin Zhang, Renrui Zhang +4
Masked Autoencoders learn strong visual representations and achieve state-of-the-art results in several independent modalities, yet very few works have addressed their capabilities…