◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Erik Visser

4 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author2
  • last author2

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • cs.CL1
  • cs.MM1
  • cs.SD1
  • eess.AS1

identity via Semantic Scholar / OpenAlex

collaborators

4 papers

eess.AS2024

VC-ENHANCE: Speech Restoration with Integrated Noise Suppression and Voice Conversion

Kyungguen Byun, Jason Filos, Erik Visser +1

Noise suppression (NS) algorithms are effective in improving speech quality in many cases. However, aggressive noise suppression can damage the target speech, reducing both speech…

cs.SD2023

Highly Controllable Diffusion-based Any-to-Any Voice Conversion Model with Frame-level Prosody Feature

Kyungguen Byun, Sunkuk Moon, Erik Visser

We propose a highly controllable voice manipulation system that can perform any-to-any voice conversion (VC) and prosody modulation simultaneously. State-of-the-art VC systems can…

cs.CL2023

Parameter Efficient Audio Captioning With Faithful Guidance Using Audio-text Shared Latent Representation

Arvind Krishna Sridhar, Yinyi Guo, Erik Visser +1

There has been significant research on developing pretrained transformer architectures for multimodal-to-text generation tasks. Albeit performance improvements, such models are fre…

cs.MM2023

Detecting False Alarms and Misses in Audio Captions

Rehana Mahfuz, Yinyi Guo, Arvind Krishna Sridhar +1

Metrics to evaluate audio captions simply provide a score without much explanation regarding what may be wrong in case the score is low. Manual human intervention is needed to find…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.