collaborators
Showing cs.CLShow all

6 papers · 1 filter

cs.CL2025

Why Not Transform Chat Large Language Models to Non-English?

Xiang Geng, Ming Zhu, Jiahuan Li +14

The scarcity of non-English data limits the development of non-English large language models (LLMs). Transforming English-centric LLMs to non-English has been identified as an effe…

cs.CL2025

Alleviating Distribution Shift in Synthetic Data for Machine Translation Quality Estimation

Xiang Geng, Zhejian Lai, Jiajun Chen +2

Quality Estimation (QE) models evaluate the quality of machine translations without reference translations, serving as the reward models for the translation task. Due to the data s…

cs.CL2024

"I've Heard of You!": Generate Spoken Named Entity Recognition Data for Unseen Entities

Jiawei Yu, Xiang Geng, Yuang Li +8

Spoken named entity recognition (NER) aims to identify named entities from speech, playing an important role in speech processing. New named entities appear every day, however, ann…

cs.CL2024

From Handcrafted Features to LLMs: A Brief Survey for Machine Translation Quality Estimation

Haofei Zhao, Yilun Liu, Shimin Tao +6

Machine Translation Quality Estimation (MTQE) is the task of estimating the quality of machine-translated text in real time without the need for reference translations, which is of…

cs.CL2024

Lost in the Source Language: How Large Language Models Evaluate the Quality of Machine Translation

Xu Huang, Zhirui Zhang, Xiang Geng +3

This study investigates how Large Language Models (LLMs) leverage source and reference data in machine translation evaluation task, aiming to better understand the mechanisms behin…

cs.CL2024

MAPO: Advancing Multilingual Reasoning through Multilingual Alignment-as-Preference Optimization

Shuaijie She, Wei Zou, Shujian Huang +4

Though reasoning abilities are considered language-agnostic, existing LLMs exhibit inconsistent reasoning abilities across different languages, e.g., reasoning in the dominant lang…