3 papers
cs.CL2026
Rubric-as-Experts: Case-Specific MQM Rubrics for Translation Quality Evaluation
Weilu Xu, Yunzhi Shen, Xinye Wang +2
Large language models (LLMs) have shown strong potential in fine-grained translation quality evaluation (QE), yet existing MQM-based approaches typically rely on fixed rubric confi…
cs.SD2026
Diachronic Modeling of Tonal Coherence on the Tonnetz Across Classical and Popular Repertoires
Weilun Xu, Edward Hall, Martin Rohrmeier
How do different musical traditions achieve tonal coherence? Most computational measures to date have analysed tonal coherence in terms of a single dimension, whereas a multi-dimen…
cs.CL2026
Probing Ethical Framework Representations in Large Language Models: Structure, Entanglement, and Methodological Challenges
Weilun Xu, Alexander Rusnak, Frederic Kaplan
When large language models make ethical judgments, do their internal representations distinguish between normative frameworks, or collapse ethics into a single acceptability dimens…