9 citations · 9 across the 3 of their papers we have counts for
5 papers
AdaSpeech 3: Adaptive Text to Speech for Spontaneous Style
Yuzi Yan, Xu Tan, Bohan Li +6
While recent text to speech (TTS) models perform very well in synthesizing reading-style (e.g., audiobook) speech, it is still challenging to synthesize spontaneous-style speech (e…
Effectively Leveraging Attributes for Visual Similarity
Samarth Mishra, Zhongping Zhang, Yuan Shen +3
Measuring similarity between two images often requires performing complex reasoning along different axes (e.g., color, texture, or shape). Insights into what might be important for…
AdaSpeech 2: Adaptive Text to Speech with Untranscribed Data
Yuzi Yan, Xu Tan, Bohan Li +4
Text to speech (TTS) is widely used to synthesize personal voice for a target speaker, where a well-trained source TTS model is fine-tuned with few paired adaptation data (speech a…
Joint Extraction of Entity and Relation with Information Redundancy Elimination
Yuanhao Shen, Jungang Han
To solve the problem of redundant information and overlapping relations of the entity and relation extraction model, we propose a joint extraction model. This model can directly ex…
Can AI decrypt fashion jargon for you?
Yuan Shen, Shanduojiao Jiang, Muhammad Rizky Wellyanto +1
When people talk about fashion, they care about the underlying meaning of fashion concepts,e.g., style.For example, people ask questions like what features make this dress smart.Ho…