◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Hung-Yi Lee

14 papers hereh-index 4124 citations14 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author1
  • last author13

Across the 14 of 14 papers where every author was matched, so the position is known.

fields
  • cs.SD6
  • cs.CL4
  • eess.AS4
same name
  • Hung-yi Lee — 137 papers, h 53
  • Hung-yi Lee — 44 papers, h 11
  • Hung-yi Lee — 16 papers, h 6
  • Hung-yi Lee — 15 papers
  • Hung-yi Lee — 15 papers, h 11
  • Hung-yi Lee — 14 papers, h 10

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

collaborators
Showing cs.CLShow all

4 papers · 1 filter

cs.CL2026

Tri-PvP: Exposing Modality Bias in Omni-Modal Large Language Models through Perceptual-Propositional Evidence Conflicts

Yen-Ting Piao, Shu-Yun Chen, Chin-Hui Chu +4

Omni-modal large language models (OLLMs) jointly process vision, audio, and text, yet their modality bias under cross-modal conflict remains underexplored. Existing benchmarks conf…

cs.CL2026

SPAR-K: Scheduled Periodic Alternating Early Exit for Spoken Language Models

Hsiao-Ying Huang, Cheng-Han Chiang, Hung-yi Lee

Interleaved spoken language models (SLMs) alternately generate text and speech tokens, but decoding at full transformer depth for every step becomes costly, especially due to long…

cs.CL2025

AudioLens: A Closer Look at Auditory Attribute Perception of Large Audio-Language Models

Chih-Kai Yang, Neo Ho, Yi-Jyun Lee +1

Understanding the internal mechanisms of large audio-language models (LALMs) is crucial for interpreting their behavior and improving performance. This work presents the first in-d…

cs.CL2025

Analyzing Mitigation Strategies for Catastrophic Forgetting in End-to-End Training of Spoken Language Models

Chi-Yuan Hsiao, Ke-Han Lu, Kai-Wei Chang +3

End-to-end training of Spoken Language Models (SLMs) commonly involves adapting pre-trained text-based Large Language Models (LLMs) to the speech modality through multi-stage train…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.