◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Yifan Yang

Shanghai Jiao Tong University

29 papers hereh-index 14937 citations37 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author7
  • middle author20

Across the 27 of 29 papers where every author was matched, so the position is known.

fields
  • eess.AS19
  • cs.SD5
  • cs.CL3
  • cs.AI1
  • cs.LG1
affiliations
  • Shanghai Jiao Tong University
Homepage
same name
  • Yifan Yang — 43 papers
  • Yifan Yang — 21 papers, h 6
  • Yifan Yang — 16 papers, h 4
  • Yifan Yang — 10 papers, h 6
  • Yifan Yang — 10 papers, h 10
  • Yifan Yang — 9 papers, h 5

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20242026
most citedSLAM-LLM: A Modular, Open-Source Multimodal Large Language Model Framework and Best Practice for Speech, Language, Audio and Music Processing

3 citations · 4 across the 10 of their papers we have counts for

collaborators
Showing cs.CLShow all

3 papers · 1 filter

cs.CL2025

FELLE: Autoregressive Speech Synthesis with Token-Wise Coarse-to-Fine Flow Matching

Hui Wang, Shujie Liu, Lingwei Meng +9

To advance continuous-valued token modeling and temporal-coherence enforcement, we propose FELLE, an autoregressive model that integrates language modeling with token-wise flow mat…

cs.CL2025

Making LLMs Better Many-to-Many Speech-to-Text Translators with Curriculum Learning

Yexing Du, Youcheng Pan, Ziyang Ma +7

Multimodal Large Language Models (MLLMs) have achieved significant success in Speech-to-Text Translation (S2TT) tasks. While most existing research has focused on English-centric t…

cs.CL2025

Exploring SSL Discrete Speech Features for Zipformer-based Contextual ASR

Mingyu Cui, Yifan Yang, Jiajun Deng +7

Self-supervised learning (SSL) based discrete speech representations are highly compact and domain adaptable. In this paper, SSL discrete speech features extracted from WavLM model…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.