From the 1 of 21 linked papers with an AI index.
21 papers
LoSA: Near-Lossless Sparse Attention for Training-Free Video Diffusion Acceleration
Enhuai Liu, Yunke Wang, Yutong Wang +2
Video diffusion transformers are costly to sample: every denoising step applies self-attention over a long 3D token sequence, a quadratic cost that dominates as resolution and dura…
StellaVLA: In-Context Structured Demonstration for Generalizable Vision-Language-Action Models
Siyu Xu, Yunke Wang, Zijian Wang +6
Vision-Language-Action (VLA) models can follow instructions and manipulate objects, but their performance often collapses out of distribution (OOD), when the scene, viewpoint, or o…
Diversifying Personalized Research Ideation against AI-Induced Homogenization
Rui Xu, Yunke Wang, Linwei Tao +2
The paper introduces DivAlign, a pipeline that creates fine‑grained researcher profiles and scores AI‑generated research directions on executability, comprehensibility, and growth…
SmoothTurn: Learning to Turn Smoothly for Agile Navigation with Quadrupedal Robots
Zunzhi You, Yunke Wang, Haolan Guo +1
Quadrupedal robots show great potential for valuable real-world applications such as fire rescue and industrial inspection. Such applications often require urgency and the ability…
PARE: Pruning and Adaptive Routing for Efficient Video Generation
Yutong Wang, Yunke Wang, Tianfan Xue +4
Video Diffusion Transformers (DiTs) generate high-quality videos but demand substantial compute due to wide blocks, deep architectures, and iterative sampling. Recent methods reduc…
BEAT: Rhythm-Elastic Alignment for Agentic Music-guided Movie Trailer Generation
Yutong Wang, Yunke Wang, Xinyuan Chen +1
Automatic movie trailer generation must select shots from a full-length film and synchronize them with background music. Existing methods either relegate music alignment to post-pr…