◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Zhenhui Ye

21 papers hereh-index 171.9k citations37 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author7
  • middle author14

Across the 21 of 21 papers where every author was matched, so the position is known.

fields
  • cs.CV10
  • cs.SD6
  • cs.LG2
  • eess.AS2
  • cs.CL1

identity via Semantic Scholar / OpenAlex

activity
20202026
most citedMake-An-Audio: Text-To-Audio Generation with Prompt-Enhanced Diffusion Models

47 citations · 135 across the 20 of their papers we have counts for

collaborators
Showing eess.ASShow all

1 paper · 1 filter

eess.AS2025

MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis

Ziyue Jiang, Yi Ren, Ruiqi Li +11

While recent zero-shot text-to-speech (TTS) models have significantly improved speech quality and expressiveness, mainstream systems still suffer from issues related to speech-text…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.