1 citations · 1 across the 3 of their papers we have counts for
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
Spatio-Temporal Similarity Volume Aggregation for Open-Vocabulary Action Recognition
Yerim So, Jiyeong Kim, Jiwon Yoon +1
Recent Open-Vocabulary Action Recognition (OVAR) methods typically aggregate visual features into a global representation before computing text alignment, a process that obscures l…
cs.CV2025
Difficulty-Aware Label-Guided Denoising for Monocular 3D Object Detection
Soyul Lee, Seungmin Baek, Dongbo Min
Monocular 3D object detection is a cost-effective solution for applications like autonomous driving and robotics, but remains fundamentally ill-posed due to inherently ambiguous de…
cs.CV2024★ 1 cited
Fine-grained Background Representation for Weakly Supervised Semantic Segmentation
Xu Yin, Woobin Im, Dongbo Min +3
Generating reliable pseudo masks from image-level labels is challenging in the weakly supervised semantic segmentation (WSSS) task due to the lack of spatial information. Prevalent…