From the 1 of 14 linked papers with an AI index.
Showing cs.LGShow all
3 papers · 1 filter
cs.LG2026
GIPO: Gaussian Importance Sampling Policy Optimization
Chengxuan Lu, Zhenquan Zhang, Shukuan Wang +3
Post-training with reinforcement learning (RL) has recently shown strong promise for advancing multimodal agents beyond supervised imitation. However, RL remains limited by poor da…
cs.LG2026
AcceRL: A Distributed Asynchronous Reinforcement Learning and World Model Framework for Vision-Language-Action Models
Chengxuan Lu, Shukuan Wang, Yanjie Li +10
Reinforcement learning (RL) for large-scale Vision-Language-Action (VLA) models is severely bottlenecked by synchronization barriers and the high cost of environment data acquisiti…
cs.LG2025
PCaM: A Progressive Focus Attention-Based Information Fusion Method for Improving Vision Transformer Domain Adaptation
Zelin Zang, Fei Wang, Liangyu Li +4
Unsupervised Domain Adaptation (UDA) aims to transfer knowledge from a labeled source domain to an unlabeled target domain. Recent UDA methods based on Vision Transformers (ViTs) h…