activity
20242026
collaborators

5 papers

cs.CL2026

VIDA: A dataset for Visually Dependent Ambiguity in Multimodal Machine Translation

Jingheng Pan, Xintong Wang, Longyue Wang +3

Ambiguity resolution is a key challenge in multimodal machine translation (MMT), where models must genuinely leverage visual input to map an ambiguous expression to its intended me…

cs.CV2025

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation

Xintong Wang, Jingheng Pan, Yixiao Liu +8

Vision-Language Translation (VLT) is a challenging task that requires accurately recognizing multilingual text embedded in images and translating it into the target language with t…

cs.CL2025

CogSteer: Cognition-Inspired Selective Layer Intervention for Efficiently Steering Large Language Models

Xintong Wang, Jingheng Pan, Liang Ding +4

Large Language Models (LLMs) achieve remarkable performance through pretraining on extensive data. This enables efficient adaptation to diverse downstream tasks. However, the lack…

cs.CL2025

Chinese Toxic Language Mitigation via Sentiment Polarity Consistent Rewrites

Xintong Wang, Yixiao Liu, Jingheng Pan +3

Detoxifying offensive language while preserving the speaker's original intent is a challenging yet critical goal for improving the quality of online interactions. Although large la…

cs.CV2024

Mitigating Hallucinations in Large Vision-Language Models with Instruction Contrastive Decoding

Xintong Wang, Jingheng Pan, Liang Ding +1

Large Vision-Language Models (LVLMs) are increasingly adept at generating contextually detailed and coherent responses from visual inputs. However, their application in multimodal…