3 papers
cs.CV2025
MoGIC: Boosting Motion Generation via Intention Understanding and Visual Context
Junyu Shi, Yong Sun, Zhiyuan Zhang +4
Existing text-driven motion generation methods often treat synthesis as a bidirectional mapping between language and motion, but remain limited in capturing the causal logic of act…
cs.CV2025
Time-Lapse Video-Based Embryo Grading via Complementary Spatial-Temporal Pattern Mining
Yong Sun, Yipeng Wang, Junyu Shi +5
Artificial intelligence has recently shown promise in automated embryo selection for In-Vitro Fertilization (IVF). However, current approaches either address partial embryo evaluat…
cs.RO2025
RoboAct-CLIP: Video-Driven Pre-training of Atomic Action Understanding for Robotics
Zhiyuan Zhang, Yuxin He, Yong Sun +3
Visual Language Models (VLMs) have emerged as pivotal tools for robotic systems, enabling cross-task generalization, dynamic environmental interaction, and long-horizon planning th…