5 citations · 8 across the 8 of their papers we have counts for
8 papers
Activating Self-Attention for Multi-Scene Absolute Pose Regression
Miso Lee, Jihwan Kim, Jae-Pil Heo
Multi-scene absolute pose regression addresses the demand for fast and memory-efficient camera pose estimation across various real-world environments. Nowadays, transformer-based m…
Self-Feedback DETR for Temporal Action Detection
Jihwan Kim, Miso Lee, Jae-Pil Heo
Temporal Action Detection (TAD) is challenging but fundamental for real-world video applications. Recently, DETR-based models have been devised for TAD but have not performed well…
When Crowd Meets Persona: Creating a Large-Scale Open-Domain Persona Dialogue Corpus
Won Ik Cho, Yoon Kyung Lee, Seoyeon Bae +5
Building a natural language dataset requires caution since word semantics is vulnerable to subtle text change or the definition of the annotated concept. Such a tendency can be see…
Tailoring Self-Supervision for Supervised Learning
WonJun Moon, Ji-Hwan Kim, Jae-Pil Heo
Recently, it is shown that deploying a proper self-supervision is a prospective way to enhance the performance of supervised learning. Yet, the benefits of self-supervision are not…
Bootstrap Equilibrium and Probabilistic Speaker Representation Learning for Self-supervised Speaker Verification
Sung Hwan Mun, Min Hyun Han, Dongjune Lee +2
In this paper, we propose self-supervised speaker representation learning strategies, which comprise of a bootstrap equilibrium speaker representation learning in the front-end and…
Self-Attentive Multi-Layer Aggregation with Feature Recalibration and Normalization for End-to-End Speaker Verification System
Soonshin Seo, Ji-Hwan Kim
One of the most important parts of an end-to-end speaker verification system is the speaker embedding generation. In our previous paper, we reported that shortcut connections-based…