18 citations · 25 across the 28 of their papers we have counts for
21 papers · 1 filter
Recurrence Is Not Enough: Causally Validating Multilingual SAE Translation Features in Gemma 2 and 3
Giang Son Nguyen, Nhi Ngoc-Yen Nguyen, Wray Buntine +1
Sparse autoencoder (SAE) features are increasingly used to explain and steer language-model behavior, but it remains unclear whether a feature found in one language context plays t…
Beyond Surface Forms: Symbolic Edits as a Test for Logical Reasoning with LLMs
Ramya Keerthy Thatikonda, Wray Buntine, Ehsan Shareghi
Logical reasoning with large language models (LLMs) is a critical capability, as it reflects a system's ability to correctly deduce hypotheses from a given context using faithful d…
En-ViMedNER: An English-Vietnamese Parallel Biomedical Corpus with UMLS Semantic Type Annotations
Nhu Vo, Phuong Nguyen, Nu Uyen Phuong Le +4
Biomedical Named Entity Recognition (NER) is fundamental to healthcare AI applications, including clinical decision support and medical information extraction. While corpora with U…
Contrastive Training with LLM-generated Near-Misses for Robust Code-Switching Speech Recognition
Tung X. Nguyen, Hieu Minh Truong, Giang Son Nguyen +3
Code-switching (CS), the alternation between multiple languages within a single utterance, remains challenging for Automatic Speech Recognition (ASR). To address this issue, we pro…
PiDA: Phonetically-Informed Data Augmentation for Robust Vietnamese Speech Translation
Giang Son Nguyen, Tung X. Nguyen, Hieu Minh Truong +3
Cascaded speech translation (ST) systems suffer from error propagation when Automatic Speech Recognition (ASR) outputs incorrect transcripts. We present the first systematic catego…
PRISM: A Multi-Dimensional Benchmark for Evaluating LLM Peer Reviewers
Ngoc Phan Phuoc Loc, Toan Huynh La Viet, Thanh Tran Khanh +8
The rapid growth in submissions to machine learning venues has strained the scientific peer-review system and intensified interest in LLM-based automated peer reviewers. However, h…