activity
20162023
most citedLarge-Scale Machine Translation between Arabic and Hebrew: Available Corpora and Initial Results

9 citations · 27 across the 10 of their papers we have counts for

collaborators

14 papers

cs.CL20243 cited

Lookback Lens: Detecting and Mitigating Contextual Hallucinations in Large Language Models Using Only Attention Maps

Yung-Sung Chuang, Linlu Qiu, Cheng-Yu Hsieh +3

When asked to summarize articles or answer questions given a passage, large language models (LLMs) can hallucinate details and respond with unsubstantiated answers that are inaccur…

cs.SD2024

Automatic Prediction of Amyotrophic Lateral Sclerosis Progression using Longitudinal Speech Transformer

Liming Wang, Yuan Gong, Nauman Dawalatabad +7

Automatic prediction of amyotrophic lateral sclerosis (ALS) disease progression provides a more efficient and objective alternative than manual approaches. We propose ALS longitudi…

cs.CL2024

Adaptive Query Rewriting: Aligning Rewriters through Marginal Probability of Conversational Answers

Tianhua Zhang, Kun Li, Hongyin Luo +3

Query rewriting is a crucial technique for passage retrieval in open-domain conversational question answering (CQA). It decontexualizes conversational queries into self-contained q…

eess.AS2024

Revisiting Self-supervised Learning of Speech Representation from a Mutual Information Perspective

Alexander H. Liu, Sung-Lin Yeh, James Glass

Existing studies on self-supervised speech representation learning have focused on developing new training methods and applying pre-trained models for different applications. Howev…

cs.CL2023

Audio-Visual Neural Syntax Acquisition

Cheng-I Jeff Lai, Freda Shi, Puyuan Peng +10

We study phrase structure induction from visually-grounded speech. The core idea is to first segment the speech waveform into sequences of word segments, and subsequently induce ph…

cs.CL2023

Comparison of Multilingual Self-Supervised and Weakly-Supervised Speech Pre-Training for Adaptation to Unseen Languages

Andrew Rouditchenko, Sameer Khurana, Samuel Thomas +6

Recent models such as XLS-R and Whisper have made multilingual speech technologies more accessible by pre-training on audio from around 100 spoken languages each. However, there ar…