1 citations · 1 across the 4 of their papers we have counts for
4 papers
From Utterance to Vividity: Training Expressive Subtitle Translation LLM via Adaptive Local Preference Optimization
Chaoqun Cui, Shijing Wang, Liangbin Huang +4
The rapid development of Large Language Models (LLMs) has significantly enhanced the general capabilities of machine translation. However, as application scenarios become more comp…
Hermes the Polyglot: A Unified Framework to Enhance Expressiveness for Multimodal Interlingual Subtitling
Chaoqun Cui, Shijing Wang, Liangbin Huang +4
Interlingual subtitling, which translates subtitles of visual media into a target language, is essential for entertainment localization but has not yet been explored in machine tra…
VL4Gaze: Unleashing Vision-Language Models for Gaze Following
Shijing Wang, Chaoqun Cui, Yaping Huang +2
Human gaze provides essential cues for interpreting attention, intention, and social interaction in visual scenes, yet gaze understanding remains largely unexplored in current visi…
Fine-grained Video Dubbing Duration Alignment with Segment Supervised Preference Optimization
Chaoqun Cui, Liangbin Huang, Shijing Wang +4
Video dubbing aims to translate original speech in visual media programs from the source language to the target language, relying on neural machine translation and text-to-speech t…