13 citations · 14 across the 4 of their papers we have counts for
4 papers
A Hybrid System of Sound Event Detection Transformer and Frame-wise Model for DCASE 2022 Task 4
Yiming Li, Zhifang Guo, Zhirong Ye +6
In this paper, we describe in detail our system for DCASE 2022 Task4. The system combines two considerably different models: an end-to-end Sound Event Detection Transformer (SEDT)…
AMD-DBSCAN: An Adaptive Multi-density DBSCAN for datasets of extremely variable density
Ziqing Wang, Zhirong Ye, Yuyang Du +4
DBSCAN has been widely used in density-based clustering algorithms. However, with the increasing demand for Multi-density clustering, previous traditional DSBCAN can not have good…
Sound Event Detection Transformer: An Event-based End-to-End Model for Sound Event Detection
Zhirong Ye, Xiangdong Wang, Hong Liu +4
Sound event detection (SED) has gained increasing attention with its wide application in surveillance, video indexing, etc. Existing models in SED mainly generate frame-level predi…
SP-SEDT: Self-supervised Pre-training for Sound Event Detection Transformer
Zhirong Ye, Xiangdong Wang, Hong Liu +4
Recently, an event-based end-to-end model (SEDT) has been proposed for sound event detection (SED) and achieves competitive performance. However, compared with the frame-based mode…