10 citations · 10 across the 5 of their papers we have counts for
5 papers · 1 filter
Local Action-Guided Motion Diffusion Model for Text-to-Motion Generation
Peng Jin, Hao Li, Zesen Cheng +6
Text-to-motion generation requires not only grounding local actions in language but also seamlessly blending these individual actions to synthesize diverse and realistic global mot…
Instance Brownian Bridge as Texts for Open-vocabulary Video Instance Segmentation
Zesen Cheng, Kehan Li, Hao Li +5
Temporally locating objects with arbitrary class texts is the primary pursuit of open-vocabulary Video Instance Segmentation (VIS). Because of the insufficient vocabulary of video…
Changes-Aware Transformer: Learning Generalized Changes Representation
Dan Wang, Licheng Jiao, Jie Chen +2
Difference features obtained by comparing the images of two periods play an indispensable role in the change detection (CD) task. However, a pair of bi-temporal images can exhibit…
Cooperative Colorization: Exploring Latent Cross-Domain Priors for NIR Image Spectrum Translation
Xingxing Yang, Jie Chen, Zaifeng Yang
Near-infrared (NIR) image spectrum translation is a challenging problem with many promising applications. Existing methods struggle with the mapping ambiguity between the NIR and t…
Dynamic Video Frame Interpolation with integrated Difficulty Pre-Assessment
Ban Chen, Xin Jin, Youxin Chen +4
Video frame interpolation(VFI) has witnessed great progress in recent years. While existing VFI models still struggle to achieve a good trade-off between accuracy and efficiency: f…