multimodal reinforcement learning 1policy alignment 1reward shaping 1visual grounding 1visual intervention 1
From the 1 of 3 linked papers with an AI index.
3 papers
cs.CV2026
SIVA-RL: Sensitivity-Invariance Visual Alignment for Multimodal Reinforcement Learning
Cheng Tang, Junzhi Ning, Min Cen +9
The paper presents SIVA-RL, a framework that uses sample-wise visual interventions to align sensitivity and invariance in multimodal reinforcement learning models, leading to bette…
cs.RO2025
GraphCoT-VLA: A 3D Spatial-Aware Reasoning Vision-Language-Action Model for Robotic Manipulation with Ambiguous Instructions
Helong Huang, Min Cen, Kai Tan +3
Vision-language-action models have emerged as a crucial paradigm in robotic manipulation. However, existing VLA models exhibit notable limitations in handling ambiguous language in…
cs.CL2025
Constraint Multi-class Positive and Unlabeled Learning for Distantly Supervised Named Entity Recognition
Yuzhe Zhang, Min Cen, Hong Zhang
Distantly supervised named entity recognition (DS-NER) has been proposed to exploit the automatically labeled training data by external knowledge bases instead of human annotations…