activity
20202025
most citedMultimodal Knowledge Alignment with Reinforcement Learning

18 citations · 50 across the 12 of their papers we have counts for

collaborators
Showing cs.CVShow all

5 papers · 1 filter

cs.CV202547 cited

IB-GAN: Disentangled Representation Learning with Information Bottleneck Generative Adversarial Networks

Insu Jeon, Wonkwang Lee, Myeongjang Pyeon +1

We propose a new GAN-based unsupervised model for disentangled representation learning. The new model is discovered in an attempt to utilize the Information Bottleneck (IB) framewo…

cs.CV2021

Pano-AVQA: Grounded Audio-Visual Question Answering on 360 Videos

Heeseung Yun, Youngjae Yu, Wonsuk Yang +2

360 videos convey holistic views for the surroundings of a scene. It provides audio-visual cues beyond pre-determined normal field of views and displays distinctive spatial…

cs.CV20212 cited

Cycled Compositional Learning between Images and Text

Jongseok Kim, Youngjae Yu, Seunghwan Lee +1

We present an approach named the Cycled Composition Network that can measure the semantic distance of the composition of image-text embedding. First, the Composition Network transi…

cs.CV2021

ACAV100M: Automatic Curation of Large-Scale Datasets for Audio-Visual Video Representation Learning

Sangho Lee, Jiwan Chung, Youngjae Yu +4

The natural association between visual observations and their corresponding sound provides powerful self-supervisory signals for learning video representations, which makes the eve…

cs.CV2020

LID 2020: The Learning from Imperfect Data Challenge Results

Yunchao Wei, Shuai Zheng, Ming-Ming Cheng +32

Learning from imperfect data becomes an issue in many industrial applications after the research community has made profound progress in supervised learning from perfectly annotate…