activity
20202022
most citedSpeaking Speed Control of End-to-End Speech Synthesis using Sentence-Level Conditioning

1 citations · 1 across the 4 of their papers we have counts for

collaborators
Showing eess.ASShow all

5 papers · 1 filter

eess.AS20242 cited

Speech Boosting: Low-Latency Live Speech Enhancement for TWS Earbuds

Hanbin Bae, Pavel Andreev, Azat Saginbaev +4

This paper introduces a speech enhancement solution tailored for true wireless stereo (TWS) earbuds on-device usage. The solution was specifically designed to support conversations…

eess.AS2022

Enhancement of Pitch Controllability using Timbre-Preserving Pitch Augmentation in FastPitch

Hanbin Bae, Young-Sun Joo

The recently developed pitch-controllable text-to-speech (TTS) model, i.e. FastPitch, was conditioned for the pitch contours. However, the quality of the synthesized speech degrade…

eess.AS2021

FastPitchFormant: Source-filter based Decomposed Modeling for Speech Synthesis

Taejun Bak, Jae-Sung Bae, Hanbin Bae +2

Methods for modeling and controlling prosody with acoustic features have been proposed for neural text-to-speech (TTS) models. Prosodic speech can be generated by conditioning acou…

eess.AS2021

A Neural Text-to-Speech Model Utilizing Broadcast Data Mixed with Background Music

Hanbin Bae, Jae-Sung Bae, Young-Sun Joo +2

Recently, it has become easier to obtain speech data from various media such as the internet or YouTube, but directly utilizing them to train a neural text-to-speech (TTS) model is…

eess.AS20201 cited

Speaking Speed Control of End-to-End Speech Synthesis using Sentence-Level Conditioning

Jae-Sung Bae, Hanbin Bae, Young-Sun Joo +3

This paper proposes a controllable end-to-end text-to-speech (TTS) system to control the speaking speed (speed-controllable TTS; SCTTS) of synthesized speech with sentence-level sp…