collaborators

5 papers

cs.CL2026

Text Style Transfer with Parameter-efficient LLM Finetuning and Round-trip Translation

Ruoxi Liu, Philipp Koehn

This paper proposes a novel method for Text Style Transfer (TST) based on parameter-efficient fine-tuning of Large Language Models (LLMs). Addressing the scarcity of parallel corpo…

eess.AS2026

Objective Evaluation of Prosody and Intelligibility in Speech Synthesis via Conditional Prediction of Discrete Tokens

Ismail Rasim Ulgen, Zongyang Du, Junchen Lu +2

Objective evaluation of synthesized speech is critical for advancing speech generation systems, yet existing metrics for intelligibility and prosody remain limited in scope and wea…

cs.CV2025

HadaSmileNet: Hadamard fusion of handcrafted and deep-learning features for enhancing facial emotion recognition of genuine smiles

Mohammad Junayed Hasan, Nabeel Mohammed, Shafin Rahman +1

The distinction between genuine and posed emotions represents a fundamental pattern recognition challenge with significant implications for data mining applications in social scien…

cs.CL2025

Speech Vecalign: an Embedding-based Method for Aligning Parallel Speech Documents

Chutong Meng, Philipp Koehn

We present Speech Vecalign, a parallel speech document alignment method that monotonically aligns speech segment embeddings and does not depend on text transcriptions. Compared to…

cs.CL2025

HiMATE: A Hierarchical Multi-Agent Framework for Machine Translation Evaluation

Shijie Zhang, Renhao Li, Songsheng Wang +3

The advancement of Large Language Models (LLMs) enables flexible and interpretable automatic evaluations. In the field of machine translation evaluation, utilizing LLMs with transl…