2 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.RO2026
OmniVLA-RL: A Vision-Language-Action Model with Spatial Understanding and Online RL
Haoxiang Jie, Yaoyuan Yan, Xiangyu Wei +4
Visual-Language-Action (VLA) models represent a paradigm shift in embodied AI, yet existing frameworks often struggle with imprecise spatial perception, suboptimal multimodal fusio…
cs.CV2023★ 2 cited
DynStatF: An Efficient Feature Fusion Strategy for LiDAR 3D Object Detection
Yao Rong, Xiangyu Wei, Tianwei Lin +2
Augmenting LiDAR input with multiple previous frames provides richer semantic information and thus boosts performance in 3D object detection, However, crowded point clouds in multi…