7 papers
REASON: Probability map-guided dual-branch fusion framework for gastric content assessment
Nu-Fnag Xiao, De-Xing Huang, Le-Tian Wang +8
Accurate assessment of gastric content from ultrasound is critical for stratifying aspiration risk at induction of general anesthesia. However, traditional methods rely on manual t…
VLA Model Post-Training via Action-Chunked PPO and Self Behavior Cloning
Si-Cheng Wang, Tian-Yu Xiang, Xiao-Hu Zhou +6
Reinforcement learning (RL) is a promising avenue for post-training vision-language-action (VLA) models, but practical deployment is hindered by sparse rewards and unstable trainin…
Online Adaptation via Dual-Stage Alignment and Self-Supervision for Fast-Calibration Brain-Computer Interfaces
Sheng-Bin Duan, Jian-Long Hao, Tian-Yu Xiang +5
Individual differences in brain activity hinder the online application of electroencephalogram (EEG)-based brain computer interface (BCI) systems. To overcome this limitation, this…
VasoMIM: Vascular Anatomy-Aware Masked Image Modeling for Vessel Segmentation
De-Xing Huang, Xiao-Hu Zhou, Mei-Jiang Gui +7
Accurate vessel segmentation in X-ray angiograms is crucial for numerous clinical applications. However, the scarcity of annotated data presents a significant challenge, which has…
Parallels Between VLA Model Post-Training and Human Motor Learning: Progress, Challenges, and Trends
Tian-Yu Xiang, Ao-Qun Jin, Xiao-Hu Zhou +11
Vision-language-action (VLA) models extend vision-language models (VLM) by integrating action generation modules for robotic manipulation. Leveraging the strengths of VLM in vision…
VLA Model-Expert Collaboration for Bi-directional Manipulation Learning
Tian-Yu Xiang, Ao-Qun Jin, Xiao-Hu Zhou +8
The emergence of vision-language-action (VLA) models has given rise to foundation models for robot manipulation. Although these models have achieved significant improvements, their…