3 papers
cs.CV2026
V-CORE: Temporally Consistent Video Understanding for Video-LLM
Zhengjian Kang, Qi Chen, Rui Liu +4
Recent Video Large Language Models (Video-LLMs) have shown strong multimodal reasoning capabilities, yet remain challenged by video understanding tasks that require consistent temp…
cs.RO2024
DRAL: Deep Reinforcement Adaptive Learning for Multi-UAVs Navigation in Unknown Indoor Environment
Kangtong Mo, Linyue Chu, Xingyu Zhang +4
Autonomous indoor navigation of UAVs presents numerous challenges, primarily due to the limited precision of GPS in enclosed environments. Additionally, UAVs' limited capacity to c…
cs.RO2024
Optimized Coordination Strategy for Multi-Aerospace Systems in Pick-and-Place Tasks By Deep Neural Network
Ye Zhang, Linyue Chu, Letian Xu +3
In this paper, we present an advanced strategy for the coordinated control of a multi-agent aerospace system, utilizing Deep Neural Networks (DNNs) within a reinforcement learning…