2 papers
cs.DC2026
Venus: An Efficient Edge Memory-and-Retrieval System for VLM-based Online Video Understanding
Shengyuan Ye, Bei Ouyang, Tianyi Qian +5
Vision-language models (VLMs) have demonstrated impressive multimodal comprehension capabilities and are being deployed in an increasing number of online video understanding applic…
cs.RO2025
CoDrone: Autonomous Drone Navigation Assisted by Edge and Cloud Foundation Models
Pengyu Chen, Tao Ouyang, Ke Luo +2
Autonomous navigation for Unmanned Aerial Vehicles faces key challenges from limited onboard computational resources, which restrict deployed deep neural networks to shallow archit…