2 papers
cs.CV2026
Not all tokens contribute equally to diffusion learning
Guoqing Zhang, Lu Shi, Wanru Xu +4
With the rapid development of conditional diffusion models, significant progress has been made in text-to-video generation. However, we observe that these models often neglect sema…
cs.CV2025
MamFusion: Multi-Mamba with Temporal Fusion for Partially Relevant Video Retrieval
Xinru Ying, Jiaqi Mo, Jingyang Lin +3
Partially Relevant Video Retrieval (PRVR) is a challenging task in the domain of multimedia retrieval. It is designed to identify and retrieve untrimmed videos that are partially r…