collaborators

10 papers

cs.LG2026

CROP: Conservative Reward for Model-based Offline Policy Optimization

Hao Li, Xiao-Hu Zhou, Shu-Hai Li +6

Offline reinforcement learning (RL) aims to optimize a policy using collected data without online interactions. Model-based approaches are particularly appealing for addressing off…

cs.CV2026

Vascular anatomy-aware self-supervised pre-training for X-ray angiogram analysis

De-Xing Huang, Chaohui Yu, Xiao-Hu Zhou +8

X-ray angiography is the gold standard imaging modality for cardiovascular diseases. However, current deep learning approaches for X-ray angiogram analysis are severely constrained…

cs.RO2026

Parallels Between VLA Model Post-Training and Human Motor Learning: Progress, Challenges, and Trends

Tian-Yu Xiang, Ao-Qun Jin, Xiao-Hu Zhou +11

Vision-language-action (VLA) models extend vision-language models (VLM) by integrating action generation modules for robotic manipulation. Leveraging the strengths of VLM in vision…

cs.CV2025

VasoMIM: Vascular Anatomy-Aware Masked Image Modeling for Vessel Segmentation

De-Xing Huang, Xiao-Hu Zhou, Mei-Jiang Gui +7

Accurate vessel segmentation in X-ray angiograms is crucial for numerous clinical applications. However, the scarcity of annotated data presents a significant challenge, which has…

cs.CV2025

REASON: Probability map-guided dual-branch fusion framework for gastric content assessment

Nu-Fnag Xiao, De-Xing Huang, Le-Tian Wang +8

Accurate assessment of gastric content from ultrasound is critical for stratifying aspiration risk at induction of general anesthesia. However, traditional methods rely on manual t…

cs.RO2025

VLA Model Post-Training via Action-Chunked PPO and Self Behavior Cloning

Si-Cheng Wang, Tian-Yu Xiang, Xiao-Hu Zhou +6

Reinforcement learning (RL) is a promising avenue for post-training vision-language-action (VLA) models, but practical deployment is hindered by sparse rewards and unstable trainin…