3 papers
cs.CL2025
Sentiment-enhanced Graph-based Sarcasm Explanation in Dialogue
Kun Ouyang, Liqiang Jing, Xuemeng Song +3
Sarcasm Explanation in Dialogue (SED) is a new yet challenging task, which aims to generate a natural language explanation for the given sarcastic dialogue that involves multiple m…
cs.MM2024
Self-Training Boosted Multi-Factor Matching Network for Composed Image Retrieval
Haokun Wen, Xuemeng Song, Jianhua Yin +3
The composed image retrieval (CIR) task aims to retrieve the desired target image for a given multimodal query, i.e., a reference image with its corresponding modification text. Th…
cs.CL2024
Multimodal Dialog Systems with Dual Knowledge-enhanced Generative Pretrained Language Model
Xiaolin Chen, Xuemeng Song, Liqiang Jing +3
Text response generation for multimodal task-oriented dialog systems, which aims to generate the proper text response given the multimodal context, is an essential yet challenging…