5 citations · 6 across the 3 of their papers we have counts for
3 papers
cs.CV2023★ 1 cited
Joint Gaze-Location and Gaze-Object Detection
Danyang Tu, Wei Shen, Wei Sun +2
This paper proposes an efficient and effective method for joint gaze location detection (GL-D) and gaze object detection (GO-D), \emph{i.e.}, gaze following detection. Current appr…
cs.CV2023
Agglomerative Transformer for Human-Object Interaction Detection
Danyang Tu, Wei Sun, Guangtao Zhai +1
We propose an agglomerative Transformer (AGER) that enables Transformer-based human-object interaction (HOI) detectors to flexibly exploit extra instance-level cues in a single-sta…
cs.CV2023★ 5 cited
Masked Autoencoders as Image Processors
Huiyu Duan, Wei Shen, Xiongkuo Min +4
Transformers have shown significant effectiveness for various vision tasks including both high-level vision and low-level vision. Recently, masked autoencoders (MAE) for feature pr…