1 citations · 2 across the 2 of their papers we have counts for
5 papers · 1 filter
Joint Gaze-Location and Gaze-Object Detection
Danyang Tu, Wei Shen, Wei Sun +2
This paper proposes an efficient and effective method for joint gaze location detection (GL-D) and gaze object detection (GO-D), \emph{i.e.}, gaze following detection. Current appr…
Agglomerative Transformer for Human-Object Interaction Detection
Danyang Tu, Wei Sun, Guangtao Zhai +1
We propose an agglomerative Transformer (AGER) that enables Transformer-based human-object interaction (HOI) detectors to flexibly exploit extra instance-level cues in a single-sta…
MD-VQA: Multi-Dimensional Quality Assessment for UGC Live Videos
Zicheng Zhang, Wei Wu, Wei Sun +5
User-generated content (UGC) live videos are often bothered by various distortions during capture procedures and thus exhibit diverse visual qualities. Such source videos are furth…
Masked Autoencoders as Image Processors
Huiyu Duan, Wei Shen, Xiongkuo Min +4
Transformers have shown significant effectiveness for various vision tasks including both high-level vision and low-level vision. Recently, masked autoencoders (MAE) for feature pr…
End-to-End Human-Gaze-Target Detection with Transformers
Danyang Tu, Xiongkuo Min, Huiyu Duan +3
In this paper, we propose an effective and efficient method for Human-Gaze-Target (HGT) detection, i.e., gaze following. Current approaches decouple the HGT detection task into sep…