4 papers
SiniticMTError: A Machine Translation Dataset with Error Annotations for Sinitic Languages
Hannah Liu, Junghyun Min, En-Shiun Annie Lee +12
Despite major advances in machine translation (MT) in recent years, progress remains limited for many low-resource languages that lack large-scale training data and linguistic reso…
mR3: Multilingual Rubric-Agnostic Reward Reasoning Models
David Anugraha, Shou-Yi Hung, Zilu Tang +3
Evaluation using Large Language Model (LLM) judges has been widely adopted in English and shown to be effective for automatic evaluation. However, their performance does not genera…
TranslationCorrect: A Unified Framework for Machine Translation Post-Editing with Predictive Error Assistance
Syed Mekael Wasti, Shou-Yi Hung, Christopher Collins +1
Machine translation (MT) post-editing and research data collection often rely on inefficient, disconnected workflows. We introduce TranslationCorrect, an integrated framework desig…
Datasheets Aren't Enough: DataRubrics for Automated Quality Metrics and Accountability
Genta Indra Winata, David Anugraha, Emmy Liu +17
High-quality datasets are fundamental to training and evaluating machine learning models, yet their creation-especially with accurate human annotations-remains a significant challe…