1 citations · 2 across the 6 of their papers we have counts for
Showing 2024Show all
2 papers · 1 filter
cs.CV2024★ 1 cited
Roadside Monocular 3D Detection Prompted by 2D Detection
Yechi Ma, Yanan Li, Wei Hua +1
Roadside monocular 3D detection requires detecting objects of predefined classes in an RGB frame and predicting their 3D attributes, such as bird's-eye-view (BEV) locations. It has…
cs.CV2024★ 1 cited
The Neglected Tails in Vision-Language Models
Shubham Parashar, Zhiqiu Lin, Tian Liu +5
Vision-language models (VLMs) excel in zero-shot recognition but their performance varies greatly across different visual concepts. For example, although CLIP achieves impressive a…