4 citations · 4 across the 2 of their papers we have counts for
3 papers
cs.CV2025
GroundingME: Exposing the Visual Grounding Gap in MLLMs through Multi-Dimensional Evaluation
Rang Li, Lei Li, Shuhuai Ren +10
Visual grounding, localizing objects from natural language descriptions, represents a critical bridge between language and vision understanding. While multimodal large language mod…
cs.CV2023
Object-centric Cross-modal Feature Distillation for Event-based Object Detection
Lei Li, Alexander Liniger, Mario Millhaeusler +3
Event cameras are gaining popularity due to their unique properties, such as their low latency and high dynamic range. One task where these benefits can be crucial is real-time obj…
cs.CV2023★ 4 cited
CPSeg: Finer-grained Image Semantic Segmentation via Chain-of-Thought Language Prompting
Lei Li
Natural scene analysis and remote sensing imagery offer immense potential for advancements in large-scale language-guided context-aware data utilization. This potential is particul…