2 citations · 5 across the 5 of their papers we have counts for
Showing 2025Show all
2 papers · 1 filter
cs.CV2025
Edit as You See: Image-guided Video Editing via Masked Motion Modeling
Zhi-Lin Huang, Yixuan Liu, Chujun Qin +4
Recent advancements in diffusion models have significantly facilitated text-guided video editing. However, there is a relative scarcity of research on image-guided video editing, a…
cs.CL2025★ 1 cited
MSWA: Refining Local Attention with Multi-ScaleWindow Attention
Yixing Xu, Shivank Nag, Dong Li +2
Transformer-based LLMs have achieved exceptional performance across a wide range of NLP tasks. However, the standard self-attention mechanism suffers from quadratic time complexity…