From the 1 of 5 linked papers with an AI index.
5 papers
CheckVLA: Execution-Time Verification with Action-Conditioned World Model for Long-Horizon Mobile Manipulation
Yushan Liu, Peibo Sun, Xintao Chao +8
The paper introduces CheckVLA, a system that uses a frozen action‑conditioned world model to verify and intervene during long‑horizon mobile manipulation when execution deviates fr…
FreeSpec: Training-Free Long Video Generation via Singular-Spectrum Reconstruction
Fangda Chen, Shanshan Zhao, Longrong Yang +3
Video diffusion models perform well in short-video synthesis, but their training-free extension to long videos often suffers from content drift, temporal inconsistency, and over-sm…
OA-WAM: Object-Addressable World Action Model for Robust Robot Manipulation
Yushan Liu, Peibo Sun, Shoujie Li +7
World Action Models (WAMs) enhance Vision-Language-Action policies by jointly predicting scene evolution and robot actions, but existing methods usually represent the predicted wor…
SWGCN: Synergy Weighted Graph Convolutional Network for Multi-Behavior Recommendation
Fangda Chen, Yueyang Wang, Chaoli Lou +2
Multi-behavior recommendation paradigms have emerged to capture diverse user activities, forecasting primary conversions (e.g., purchases) by leveraging secondary signals like brow…
JointTuner: Appearance-Motion Adaptive Joint Training for Customized Video Generation
Fangda Chen, Shanshan Zhao, Chuanfu Xu +1
Recent advancements in customized video generation have led to significant improvements in the simultaneous adaptation of appearance and motion. Typically, decoupling the appearanc…