1 citations · 1 across the 2 of their papers we have counts for
3 papers
cs.CV2025★ 1 cited
SO-DETR: Leveraging Dual-Domain Features and Knowledge Distillation for Small Object Detection
Huaxiang Zhang, Hao Zhang, Aoran Mei +2
Detection Transformer-based methods have achieved significant advancements in general object detection. However, challenges remain in effectively detecting small objects. One key d…
cs.RO2024
ReplanVLM: Replanning Robotic Tasks with Visual Language Models
Aoran Mei, Guo-Niu Zhu, Huaxiang Zhang +1
Large language models (LLMs) have gained increasing popularity in robotic task planning due to their exceptional abilities in text analytics and generation, as well as their broad…
cs.CV2024
InsightSee: Advancing Multi-agent Vision-Language Models for Enhanced Visual Understanding
Huaxiang Zhang, Yaojia Mu, Guo-Niu Zhu +1
Accurate visual understanding is imperative for advancing autonomous systems and intelligent robots. Despite the powerful capabilities of vision-language models (VLMs) in processin…