activity
20182023
most citedClova Baseline System for the VoxCeleb Speaker Recognition Challenge 2020

97 citations · 279 across the 16 of their papers we have counts for

collaborators
Showing cs.CVShow all

14 papers · 1 filter

cs.CV20231 cited

TalkNCE: Improving Active Speaker Detection with Talk-Aware Contrastive Learning

Chaeyoung Jung, Suyeon Lee, Kihyun Nam +4

The goal of this work is Active Speaker Detection (ASD), a task to determine whether a person is speaking or not in a series of video frames. Previous works have dealt with the tas…

cs.CV20232 cited

SlowFast Network for Continuous Sign Language Recognition

Junseok Ahn, Youngjoon Jang, Joon Son Chung

The objective of this work is the effective extraction of spatial and dynamic features for Continuous Sign Language Recognition (CSLR). To accomplish this, we utilise a two-pathway…

cs.CV2023

Sound Source Localization is All about Cross-Modal Alignment

Arda Senocak, Hyeonggon Ryu, Junsik Kim +3

Humans can easily perceive the direction of sound sources in a visual scene, termed sound source localization. Recent studies on learning-based sound source localization have mainl…

cs.CV2023

That's What I Said: Fully-Controllable Talking Face Generation

Youngjoon Jang, Kyeongha Rho, Jong-Bin Woo +5

The goal of this paper is to synthesise talking faces with controllable facial motions. To achieve this goal, we propose two key ideas. The first is to establish a canonical space…

cs.CV20221 cited

MarginNCE: Robust Sound Localization with a Negative Margin

Sooyoung Park, Arda Senocak, Joon Son Chung

The goal of this work is to localize sound sources in visual scenes with a self-supervised approach. Contrastive learning in the context of sound source localization leverages the…

cs.CV20226 cited

Signing Outside the Studio: Benchmarking Background Robustness for Continuous Sign Language Recognition

Youngjoon Jang, Youngtaek Oh, Jae Won Cho +3

The goal of this work is background-robust continuous sign language recognition. Most existing Continuous Sign Language Recognition (CSLR) benchmarks have fixed backgrounds and are…