1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.CV2026
Enhancing Open-Vocabulary Object Detection through Multi-Level Fine-Grained Visual-Language Alignment
Tianyi Zhang, Antoine Simoulin, Kai Li +5
Traditional object detection systems are typically constrained to predefined categories, limiting their applicability in dynamic environments. In contrast, open-vocabulary object d…
cs.CV2024★ 1 cited
APLe: Token-Wise Adaptive for Multi-Modal Prompt Learning
Guiming Cao, Kaize Shi, Hong Fu +2
Pre-trained Vision-Language (V-L) models set the benchmark for generalization to downstream tasks among the noteworthy contenders. Many characteristics of the V-L model have been e…