collaborators

19 papers

cs.CL2026

Contrastive Training with LLM-generated Near-Misses for Robust Code-Switching Speech Recognition

Tung X. Nguyen, Hieu Minh Truong, Giang Son Nguyen +3

Code-switching (CS), the alternation between multiple languages within a single utterance, remains challenging for Automatic Speech Recognition (ASR). To address this issue, we pro…

cs.CL2026

ViMedCSS: A Vietnamese Medical Code-Switching Speech Dataset & Benchmark

Tung X. Nguyen, Nhu Vo, Giang-Son Nguyen +6

Code-switching (CS), which is when Vietnamese speech uses English words like drug names or procedures, is a common phenomenon in Vietnamese medical communication. This creates chal…

cs.CL2026

PiDA: Phonetically-Informed Data Augmentation for Robust Vietnamese Speech Translation

Giang Son Nguyen, Tung X. Nguyen, Hieu Minh Truong +3

Cascaded speech translation (ST) systems suffer from error propagation when Automatic Speech Recognition (ASR) outputs incorrect transcripts. We present the first systematic catego…

cs.CL2026

PRISM: A Multi-Dimensional Benchmark for Evaluating LLM Peer Reviewers

Ngoc Phan Phuoc Loc, Toan Huynh La Viet, Thanh Tran Khanh +8

The rapid growth in submissions to machine learning venues has strained the scientific peer-review system and intensified interest in LLM-based automated peer reviewers. However, h…

cs.AI2026

Position: agentic AI orchestration should be Bayes-consistent

Theodore Papamarkou, Pierre Alquier, Matthias Bauer +27

LLMs excel at predictive tasks and complex reasoning tasks, but many high-value deployments rely on decisions under uncertainty, for example, which tool to call, which expert to co…

cs.CL2026

Improving Symbolic Translation of Language Models for Logical Reasoning

Ramya Keerthy Thatikonda, Jiuzhou Han, Wray Buntine +1

The use of formal language for deductive logical reasoning aligns well with language models (LMs), where translating natural language (NL) into first-order logic (FOL) and employin…