5 papers
SkillReason: Reasoning-Enhanced Agent Skill Retrieval for Implicit User Requests
Donghong Jiang, Endian Lin, Luoping Cui +6
Large language model agents increasingly rely on reusable skills to extend their capabilities beyond parametric knowl- edge. However, retrieving the appropriate skill from a large-…
DSAA: Dual-Stage Attribute Activation for Fine-grained Open Vocabulary Detection
Donghong Jiang, Endian Lin, Hanqing Liu +4
Open-Vocabulary Object Detection (OVD) models break the limitations of closed-set detection, enabling the identification of unseen categories through natural language prompts. Howe…
RE-VLM: Event-Augmented Vision-Language Model for Scene Understanding
Hanqing Liu, Mingjie Liu, Luoping Cui +3
Conventional vision-language models (VLMs) struggle to interpret scenes captured under adverse conditions (e.g., low light, high dynamic range, or fast motion) because standard RGB…
Enhancing Event-based Object Detection with Monocular Normal Maps
Mingjie Liu, Hanqing Liu, Luoping Cui +1
Object detection in autonomous driving is frequently compromised by complex illumination. While event cameras offer a robust solution, they are susceptible to sudden contrast chang…
PEOD: A Pixel-Aligned Event-RGB Benchmark for Object Detection under Challenging Conditions
Luoping Cui, Hanqing Liu, Mingjie Liu +4
Robust object detection for challenging scenarios increasingly relies on event cameras, yet existing Event-RGB datasets remain constrained by sparse coverage of extreme conditions…