13 citations · 15 across the 14 of their papers we have counts for
Showing 2024Show all
3 papers · 1 filter
cs.CV2024★ 2 cited
VisionGRU: A Linear-Complexity RNN Model for Efficient Image Analysis
Shicheng Yin, Kaixuan Yin, Weixing Chen +2
Convolutional Neural Networks (CNNs) and Vision Transformers (ViTs) are two dominant models for image analysis. While CNNs excel at extracting multi-scale features and ViTs effecti…
cs.CV2024
Towards Long-Horizon Vision-Language Navigation: Platform, Benchmark and Method
Xinshuai Song, Weixing Chen, Yang Liu +3
Existing Vision-Language Navigation (VLN) methods primarily focus on single-stage navigation, limiting their effectiveness in multi-stage and long-horizon tasks within complex and…
cs.CV2024
Aligning Cyber Space with Physical World: A Comprehensive Survey on Embodied AI
Yang Liu, Weixing Chen, Yongjie Bai +4
Embodied Artificial Intelligence (Embodied AI) is crucial for achieving Artificial General Intelligence (AGI) and serves as a foundation for various applications (e.g., intelligent…