activity
20172022
most citedRegularizing Deep Neural Networks by Noise: Its Interpretation and Optimization

78 citations · 127 across the 7 of their papers we have counts for

collaborators
Showing cs.CVShow all

5 papers · 1 filter

cs.CV20223 cited

MSTR: Multi-Scale Transformer for End-to-End Human-Object Interaction Detection

Bumsoo Kim, Jonghwan Mun, Kyoung-Woon On +3

Human-Object Interaction (HOI) detection is the task of identifying a set of <human, object, interaction> triplets from an image. Recent work proposed transformer encoder-decoder a…

cs.CV20223 cited

Boundary-aware Self-supervised Learning for Video Scene Segmentation

Jonghwan Mun, Minchul Shin, Gunsoo Han +4

Self-supervised learning has drawn attention through its effectiveness in learning in-domain representations with no ground-truth annotations; in particular, it is shown that prope…

cs.CV20214 cited

Winning the ICCV'2021 VALUE Challenge: Task-aware Ensemble and Transfer Learning with Visual Concepts

Minchul Shin, Jonghwan Mun, Kyoung-Woon On +3

The VALUE (Video-And-Language Understanding Evaluation) benchmark is newly introduced to evaluate and analyze multi-modal representation learning algorithms on three video-and-lang…

cs.CV202027 cited

Local-Global Video-Text Interactions for Temporal Grounding

Jonghwan Mun, Minsu Cho, Bohyung Han

This paper addresses the problem of text-to-video temporal grounding, which aims to identify the time interval in a video semantically relevant to a text query. We tackle this prob…

cs.CV20197 cited

Streamlined Dense Video Captioning

Jonghwan Mun, Linjie Yang, Zhou Ren +2

Dense video captioning is an extremely challenging task since accurate and coherent description of events in a video requires holistic understanding of video contents as well as co…