1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.RO2026
OmniVLA-RL: A Vision-Language-Action Model with Spatial Understanding and Online RL
Haoxiang Jie, Yaoyuan Yan, Xiangyu Wei +4
Visual-Language-Action (VLA) models represent a paradigm shift in embodied AI, yet existing frameworks often struggle with imprecise spatial perception, suboptimal multimodal fusio…
cs.CV2025★ 1 cited
Increasing the Diversity in RGB-to-Thermal Image Translation for Automotive Applications
Kaili Wang, Leonardo Ravaglia, Roberto Longo +5
Thermal imaging in Advanced Driver Assistance Systems (ADAS) improves road safety with superior perception in low-light and harsh weather conditions compared to traditional RGB cam…