3 citations · 5 across the 3 of their papers we have counts for
3 papers
eess.AS2025
Performance improvement of spatial semantic segmentation with enriched audio features and agent-based error correction for DCASE 2025 Challenge Task 4
Jongyeon Park, Joonhee Lee, Do-Hyeon Lim +3
This technical report presents submission systems for Task 4 of the DCASE 2025 Challenge. This model incorporates additional audio features (spectral roll-off and chroma features)…
eess.AS2023★ 2 cited
GIST-AiTeR Speaker Diarization System for VoxCeleb Speaker Recognition Challenge (VoxSRC) 2023
Dongkeon Park, Ji Won Kim, Kang Ryeol Kim +2
This report describes the submission system by the GIST-AiTeR team for the VoxCeleb Speaker Recognition Challenge 2023 (VoxSRC-23) Track 4. Our submission system focuses on impleme…
eess.AS2023★ 3 cited
Semi-supervsied Learning-based Sound Event Detection using Freuqency Dynamic Convolution with Large Kernel Attention for DCASE Challenge 2023 Task 4
Ji Won Kim, Sang Won Son, Yoonah Song +3
This report proposes a frequency dynamic convolution (FDY) with a large kernel attention (LKA)-convolutional recurrent neural network (CRNN) with a pre-trained bidirectional encode…