◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Xu Tan

4 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author4

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • eess.AS2
  • cs.CL1
  • cs.SD1
same name
  • Xu Tan — 43 papers
  • Xu Tan — 17 papers, h 18
  • Xu Tan — 12 papers
  • Xu Tan — 3 papers
  • Xu Tan — 2 papers
  • Xu Tan — 2 papers

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

most citedNaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models

20 citations · 23 across the 4 of their papers we have counts for

collaborators

4 papers

eess.AS2024★ 20 cited

NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models

Zeqian Ju, Yuancheng Wang, Kai Shen +16

While recent large-scale text-to-speech (TTS) models have achieved significant progress, they still fall short in speech quality, similarity, and prosody. Considering speech intric…

cs.CL2023★ 2 cited

MusicAgent: An AI Agent for Music Understanding and Generation with Large Language Models

Dingyao Yu, Kaitao Song, Peiling Lu +5

AI-empowered music processing is a diverse field that encompasses dozens of tasks, ranging from generation tasks (e.g., timbre synthesis) to comprehension tasks (e.g., music classi…

eess.AS2023

PromptTTS 2: Describing and Generating Voices with Text Prompt

Yichong Leng, Zhifang Guo, Kai Shen +12

Speech conveys more information than text, as the same word can be uttered in various voices to convey diverse information. Compared to traditional text-to-speech (TTS) methods rel…

cs.SD2023★ 1 cited

MelodyGLM: Multi-task Pre-training for Symbolic Melody Generation

Xinda Wu, Zhijie Huang, Kejun Zhang +5

Pre-trained language models have achieved impressive results in various music understanding and generation tasks. However, existing pre-training methods for symbolic melody generat…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.