2 papers
cs.RO2026
Learning to Feel the Future: DreamTacVLA for Contact-Rich Manipulation
Guo Ye, Zexi Zhang, Xu Zhao +4
Vision-Language-Action (VLA) models have shown remarkable generalization by mapping web-scale knowledge to robotic control, yet they remain blind to physical contact. Consequently,…
cs.AI2025
FairReason: Balancing Reasoning and Social Bias in MLLMs
Zhenyu Pan, Yutong Zhang, Jianshu Zhang +6
Multimodal Large Language Models (MLLMs) already achieve state-of-the-art results across a wide range of tasks and modalities. To push their reasoning ability further, recent studi…