1 citations · 1 across the 3 of their papers we have counts for
4 papers
MoLT: Mixture of Layer-Wise Tokens for Efficient Audio-Visual Learning
Kyeongha Rho, Hyeongkeun Lee, Jae Won Cho +1
In this paper, we propose Mixture of Layer-Wise Tokens (MoLT), a parameter- and memory-efficient adaptation framework for audio-visual learning. The key idea of MoLT is to replace…
LAVCap: LLM-based Audio-Visual Captioning using Optimal Transport
Kyeongha Rho, Hyeongkeun Lee, Valentio Iverson +1
Automated audio captioning is a task that generates textual descriptions for audio content, and recent studies have explored using visual information to enhance captioning quality.…
NTIRE 2020 Challenge on Real Image Denoising: Dataset, Methods and Results
Abdelrahman Abdelhamed, Mahmoud Afifi, Radu Timofte +87
This paper reviews the NTIRE 2020 challenge on real image denoising with focus on the newly introduced dataset, the proposed methods and their results. The challenge is a new versi…
NTIRE 2020 Challenge on Spectral Reconstruction from an RGB Image
Boaz Arad, Radu Timofte, Ohad Ben-Shahar +4
This paper reviews the second challenge on spectral reconstruction from RGB images, i.e., the recovery of whole-scene hyperspectral (HS) information from a 3-channel RGB image. As…