2 citations · 5 across the 4 of their papers we have counts for
Showing eess.ASShow all
3 papers · 1 filter
eess.AS2023★ 1 cited
Knowledge Distillation for Efficient Audio-Visual Video Captioning
Özkan Çaylı, Xubo Liu, Volkan Kılıç +1
Automatically describing audio-visual content with texts, namely video captioning, has received significant attention due to its potential applications across diverse fields. Deep…
eess.AS2023
Dual Transformer Decoder based Features Fusion Network for Automated Audio Captioning
Jianyuan Sun, Xubo Liu, Xinhao Mei +3
Automated audio captioning (AAC) which generates textual descriptions of audio content. Existing AAC models achieve good results but only use the high-dimensional representation of…
eess.AS2022★ 2 cited
Deep Neural Decision Forest for Acoustic Scene Classification
Jianyuan Sun, Xubo Liu, Xinhao Mei +4
Acoustic scene classification (ASC) aims to classify an audio clip based on the characteristic of the recording environment. In this regard, deep learning based approaches have eme…