22 citations · 23 across the 3 of their papers we have counts for
3 papers
cs.CV2024
Bridging Vision and Language Spaces with Assignment Prediction
Jungin Park, Jiyoung Lee, Kwanghoon Sohn
This paper introduces VLAP, a novel approach that bridges pretrained vision models and large language models (LLMs) to make frozen LLMs understand the visual world. VLAP transforms…
cs.CV2023★ 22 cited
TemporalMaxer: Maximize Temporal Context with only Max Pooling for Temporal Action Localization
Tuan N. Tang, Kwonyoung Kim, Kwanghoon Sohn
Temporal Action Localization (TAL) is a challenging task in video understanding that aims to identify and localize actions within a video sequence. Recent studies have emphasized t…
cs.CV2016★ 1 cited
Deeply Aggregated Alternating Minimization for Image Restoration
Youngjung Kim, Hyungjoo Jung, Dongbo Min +1
Regularization-based image restoration has remained an active research topic in computer vision and image processing. It often leverages a guidance signal captured in different fie…