7 citations · 18 across the 4 of their papers we have counts for
4 papers
Audio-Visual Segmentation via Unlabeled Frame Exploitation
Jinxiang Liu, Yikun Liu, Fei Zhang +3
Audio-visual segmentation (AVS) aims to segment the sounding objects in video frames. Although great progress has been witnessed, we experimentally reveal that current methods reac…
Audio-aware Query-enhanced Transformer for Audio-Visual Segmentation
Jinxiang Liu, Chen Ju, Chaofan Ma +3
The goal of the audio-visual segmentation (AVS) task is to segment the sounding objects in the video frames using audio cues. However, current fusion-based methods have the perform…
DiffusionSeg: Adapting Diffusion Towards Unsupervised Object Discovery
Chaofan Ma, Yuhuan Yang, Chen Ju +5
Learning from a large corpus of data, pre-trained models have achieved impressive progress nowadays. As popular generative pre-training, diffusion models capture both low-level vis…
Constraint and Union for Partially-Supervised Temporal Sentence Grounding
Chen Ju, Haicheng Wang, Jinxiang Liu +5
Temporal sentence grounding aims to detect the event timestamps described by the natural language query from given untrimmed videos. The existing fully-supervised setting achieves…