3 papers
cs.LG2025
From Data-Centric to Sample-Centric: Enhancing LLM Reasoning via Progressive Optimization
Xinjie Chen, Minpeng Liao, Guoxin Chen +4
Reinforcement learning with verifiable rewards (RLVR) has recently advanced the reasoning capabilities of large language models (LLMs). While prior work has emphasized algorithmic…
cs.CL2025
LLMs Can Achieve High-quality Simultaneous Machine Translation as Efficiently as Offline
Biao Fu, Minpeng Liao, Kai Fan +4
When the complete source sentence is provided, Large Language Models (LLMs) perform excellently in offline machine translation even with a simple prompt "Translate the following se…
cs.CL2025
Efficient and Adaptive Simultaneous Speech Translation with Fully Unidirectional Architecture
Biao Fu, Donglei Yu, Minpeng Liao +4
Simultaneous speech translation (SimulST) produces translations incrementally while processing partial speech input. Although large language models (LLMs) have showcased strong cap…