3 papers
cs.CL2026
Reasoning Model Is Superior LLM-Judge, Yet Suffers from Biases
Hui Huang, Xuanxin Wu, Muyun Yang +1
This paper presents the first systematic comparison investigating whether Large Reasoning Models (LRMs) are superior judges to non-reasoning LLMs. Our empirical analysis yields fou…
cs.CL2026
Towards Fine-Grained Code-Switch Speech Translation with Semantic Space Alignment
Yan Gao, Yazheng Yang, Zhibin Lan +5
Code-switching (CS) speech translation (ST) aims to translate speech that alternates between multiple languages into a target language text, posing significant challenges due to th…
cs.CL2026
DiVA: Fine-grained Factuality Verification with Agentic-Discriminative Verifier
Hui Huang, Muyun Yang, Yuki Arase
Despite the significant advancements of Large Language Models (LLMs), their factuality remains a critical challenge, fueling growing interest in factuality verification. Existing r…