◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Qian Yang

4 papers hereh-index 5305 citations10 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author1
  • middle author2

Across the 3 of 4 papers where every author was matched, so the position is known.

fields
  • eess.AS3
  • cs.SD1
same name
  • Qian Yang — 21 papers, h 24
  • Qian Yang — 10 papers, h 4
  • Qian Yang — 9 papers, h 4
  • Qian Yang — 8 papers, h 21
  • Qian Yang — 7 papers, h 16
  • Qian Yang — 7 papers, h 14

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

most citedWavChat: A Survey of Spoken Dialogue Models

2 citations · 2 across the 3 of their papers we have counts for

collaborators

4 papers

cs.SD2025

Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model

Jialong Zuo, Shengpeng Ji, Minghui Fang +8

This paper introduces PFlow-VC, a conditional flow matching voice conversion model that leverages fine-grained discrete pitch tokens and target speaker prompt information for expre…

eess.AS2024★ 2 cited

WavChat: A Survey of Spoken Dialogue Models

Shengpeng Ji, Yifu Chen, Minghui Fang +16

Recent advancements in spoken dialogue models, exemplified by systems like GPT-4o, have captured significant attention in the speech domain. Compared to traditional three-tier casc…

eess.AS2024

WavTokenizer: an Efficient Acoustic Discrete Codec Tokenizer for Audio Language Modeling

Shengpeng Ji, Ziyue Jiang, Wen Wang +14

Language models have been effectively applied to modeling natural signals, such as images, video, speech, and audio. A crucial component of these models is the codec tokenizer, whi…

eess.AS2024

MSceneSpeech: A Multi-Scene Speech Dataset For Expressive Speech Synthesis

Qian Yang, Jialong Zuo, Zhe Su +6

We introduce an open source high-quality Mandarin TTS dataset MSceneSpeech (Multiple Scene Speech Dataset), which is intended to provide resources for expressive speech synthesis.…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.