◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Rui Liu

4 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author3
  • last author1

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • cs.CL2
  • cs.CV1
  • cs.MM1
same name
  • Rui Liu — 48 papers, h 36
  • Rui Liu — 15 papers
  • Rui Liu — 11 papers
  • Rui Liu — 6 papers, h 12
  • Rui Liu — 6 papers, h 12
  • Rui Liu — 6 papers

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

most citedRetrieval-Augmented Dialogue Knowledge Aggregation for Expressive Conversational Speech Synthesis

1 citations · 1 across the 4 of their papers we have counts for

collaborators

4 papers

cs.CV2025

Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech

Rui Liu, Shuwei He, Yifan Hu +1

Visual Text-to-Speech (VTTS) aims to take the environmental image as the prompt to synthesize the reverberant speech for the spoken content. The challenge of this task lies in unde…

cs.CL2025★ 1 cited

Retrieval-Augmented Dialogue Knowledge Aggregation for Expressive Conversational Speech Synthesis

Rui Liu, Zhenqi Jia, Feilong Bao +1

Conversational speech synthesis (CSS) aims to take the current dialogue (CD) history as a reference to synthesize expressive speech that aligns with the conversational style. Unlik…

cs.MM2025

Listening and Seeing Again: Generative Error Correction for Audio-Visual Speech Recognition

Rui Liu, Hongyu Yuan, Haizhou Li

Unlike traditional Automatic Speech Recognition (ASR), Audio-Visual Speech Recognition (AVSR) takes audio and visual signals simultaneously to infer the transcription. Recent studi…

cs.CL2024

Intra- and Inter-modal Context Interaction Modeling for Conversational Speech Synthesis

Zhenqi Jia, Rui Liu

Conversational Speech Synthesis (CSS) aims to effectively take the multimodal dialogue history (MDH) to generate speech with appropriate conversational prosody for target utterance…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.