4 papers
Bridging the Visual-to-Physical Gap: Physically Aligned Representations for Fall Risk Analysis
Xianqi Zhang
Vision-based fall analysis has advanced rapidly, but a key bottleneck remains: visually similarmotions can correspond to very different physical outcomes because small differences…
Region-Level Context-Aware Multimodal Understanding
Hongliang Wei, Xianqi Zhang, Xingtao Wang +2
Despite significant progress, existing research on Multimodal Large Language Models (MLLMs) mainly focuses on general visual understanding, overlooking the ability to integrate tex…
Task-Agnostic Learning to Accomplish New Tasks
Xianqi Zhang, Xingtao Wang, Xu Liu +3
Reinforcement Learning (RL) and Imitation Learning (IL) have made great progress in robotic decision-making in recent years. However, these methods show obvious deterioration for n…
FLAM: Foundation Model-Based Body Stabilization for Humanoid Locomotion and Manipulation
Xianqi Zhang, Hongliang Wei, Wenrui Wang +3
Humanoid robots have attracted significant attention in recent years. Reinforcement Learning (RL) is one of the main ways to control the whole body of humanoid robots. RL enables a…