5 citations · 6 across the 3 of their papers we have counts for
3 papers
cs.CV2025★ 1 cited
SO-DETR: Leveraging Dual-Domain Features and Knowledge Distillation for Small Object Detection
Huaxiang Zhang, Hao Zhang, Aoran Mei +2
Detection Transformer-based methods have achieved significant advancements in general object detection. However, challenges remain in effectively detecting small objects. One key d…
cs.RO2024
ReplanVLM: Replanning Robotic Tasks with Visual Language Models
Aoran Mei, Guo-Niu Zhu, Huaxiang Zhang +1
Large language models (LLMs) have gained increasing popularity in robotic task planning due to their exceptional abilities in text analytics and generation, as well as their broad…
eess.IV2021★ 5 cited
A New Image Codec Paradigm for Human and Machine Uses
Sien Chen, Jian Jin, Lili Meng +5
With the AI of Things (AIoT) development, a huge amount of visual data, e.g., images and videos, are produced in our daily work and life. These visual data are not only used for hu…