9 citations · 13 across the 15 of their papers we have counts for
Showing 2024 · cs.CVShow all
2 papers · 2 filters
cs.CV2024★ 2 cited
Interpreting Object-level Foundation Models via Visual Precision Search
Ruoyu Chen, Siyuan Liang, Jingzhi Li +5
Advances in multimodal pre-training have propelled object-level foundation models, such as Grounding DINO and Florence-2, in tasks like visual grounding and object detection. Howev…
cs.CV2024★ 2 cited
Object Detectors in the Open Environment: Challenges, Solutions, and Outlook
Siyuan Liang, Wei Wang, Ruoyu Chen +5
With the emergence of foundation models, deep learning-based object detectors have shown practical usability in closed set scenarios. However, for real-world tasks, object detector…