23 citations · 31 across the 9 of their papers we have counts for
10 papers · 1 filter
RFMSR: Residual Flow Matching for Image Super-Resolution
Shuwei Huang, Tianyao Luo, Jicheng Liu +1
Image super-resolution (ISR) has witnessed remarkable progress with diffusion models and flow matching. The dominant text-to-image (T2I) based approaches leverage large-scale found…
Fewer Steps, Better Performance: Efficient Cross-Modal Clip Trimming for Video Moment Retrieval Using Language
Xiang Fang, Daizong Liu, Wanlong Fang +5
Given an untrimmed video and a sentence query, video moment retrieval using language (VMR) aims to locate a target query-relevant moment. Since the untrimmed video is overlong, alm…
3DHacker: Spectrum-based Decision Boundary Generation for Hard-label 3D Point Cloud Attack
Yunbo Tao, Daizong Liu, Pan Zhou +3
With the maturity of depth sensors, the vulnerability of 3D point cloud models has received increasing attention in various applications such as autonomous driving and robot naviga…
Transform-Equivariant Consistency Learning for Temporal Sentence Grounding
Daizong Liu, Xiaoye Qu, Jianfeng Dong +6
This paper addresses the temporal sentence grounding (TSG). Although existing methods have made decent achievements in this task, they not only severely rely on abundant video-quer…
Jointly Visual- and Semantic-Aware Graph Memory Networks for Temporal Sentence Localization in Videos
Daizong Liu, Pan Zhou
Temporal sentence localization in videos (TSLV) aims to retrieve the most interested segment in an untrimmed video according to a given sentence query. However, almost of existing…
You Can Ground Earlier than See: An Effective and Efficient Pipeline for Temporal Sentence Grounding in Compressed Videos
Xiang Fang, Daizong Liu, Pan Zhou +1
Given an untrimmed video, temporal sentence grounding (TSG) aims to locate a target moment semantically according to a sentence query. Although previous respectable works have made…