◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Hung-yi Lee

6 papers hereh-index 116 citations8 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author2
  • last author4

Across the 6 of 6 papers where every author was matched, so the position is known.

fields
  • eess.AS4
  • cs.CL2
same name
  • Hung-yi Lee — 44 papers, h 11
  • Hung-yi Lee — 13 papers, h 6
  • Hung-yi Lee — 13 papers, h 7
  • Hung-yi Lee — 13 papers, h 8
  • Hung-yi Lee — 12 papers, h 4
  • Hung-yi Lee — 12 papers, h 10

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

collaborators
Showing eess.ASShow all

4 papers · 1 filter

eess.AS2026

Hearing Like Humans? Sound Symbolism and Perceptual Alignment in Speech Language Models

Yun-Shao Tsai, Chun-Wei Chen, Chee-En Yu +2

Sound symbolism, the human tendency to map speech sounds to perceptual qualities such as roundness or sharpness, arises primarily from the acoustics of speech rather than spelling.…

eess.AS2026

The False Resonance: A Critical Examination of Emotion Embedding Similarity for Speech Generation Evaluation

Yun-Shao Tsai, Yi-Cheng Lin, Huang-Cheng Chou +5

Objective metrics for emotional expressiveness are vital for speech generation, particularly in expressive synthesis and voice conversion requiring emotional prosody transfer. To q…

eess.AS2026

Walking Through Uncertainty: An Empirical Study of Uncertainty Estimation for Audio-Aware Large Language Models

Chun-Yi Kuan, Wei-Ping Huang, Hung-yi Lee

Recent audio-aware large language models (ALLMs) have demonstrated strong capabilities across diverse audio understanding and reasoning tasks, but they still frequently produce hal…

eess.AS2026

ParaS2S: Benchmarking and Aligning Spoken Language Models for Paralinguistic-aware Speech-to-Speech Interaction

Shu-wen Yang, Ming Tu, Andy T. Liu +5

Speech-to-Speech (S2S) models have shown promising dialogue capabilities, but their ability to handle paralinguistic cues - such as emotion, tone, and speaker attributes - and to r…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.