Showing cs.CLShow all
3 papers · 1 filter
cs.CL2025
XL-Suite: Cross-Lingual Synthetic Training and Evaluation Data for Open-Ended Generation
Vivek Iyer, Pinzhen Chen, Ricardo Rei +1
Cross-lingual open-ended generation - responding in a language different from that of the query - is an important yet understudied problem. This work proposes XL-Instruct, a novel…
cs.CL2025
Generalizing From Short to Long: Effective Data Synthesis for Long-Context Instruction Tuning
Wenhao Zhu, Pinzhen Chen, Hanxu Hu +4
Long-context modelling for large language models (LLMs) has been a key area of recent research because many real world use cases require reasoning over longer inputs such as docume…
cs.CL2024
Cultural Adaptation of Menus: A Fine-Grained Approach
Zhonghe Zhang, Xiaoyu He, Vivek Iyer +1
Machine Translation of Culture-Specific Items (CSIs) poses significant challenges. Recent work on CSI translation has shown some success using Large Language Models (LLMs) to adapt…