64 citations · 77 across the 3 of their papers we have counts for
3 papers
cs.CV2025★ 1 cited
TD3Net: A temporal densely connected multi-dilated convolutional network for lipreading
Byung Hoon Lee, Wooseok Shin, Sung Won Han
The word-level lipreading approach typically employs a two-stage framework with separate frontend and backend architectures to model dynamic lip movements. Each component has been…
eess.AS2023★ 64 cited
Patch-Mix Contrastive Learning with Audio Spectrogram Transformer on Respiratory Sound Classification
Sangmin Bae, June-Woo Kim, Won-Yang Cho +7
Respiratory sound contains crucial information for the early diagnosis of fatal lung diseases. Since the COVID-19 pandemic, there has been a growing interest in contact-free medica…
cs.SD2022★ 12 cited
Multi-View Attention Transfer for Efficient Speech Enhancement
Wooseok Shin, Hyun Joon Park, Jin Sob Kim +2
Recent deep learning models have achieved high performance in speech enhancement; however, it is still challenging to obtain a fast and low-complexity model without significant per…