2 citations · 7 across the 14 of their papers we have counts for
8 papers
Flickr30K-CFQ: A Compact and Fragmented Query Dataset for Text-image Retrieval
Haoyu Liu, Yaoxian Song, Xuwu Wang +4
With the explosive growth of multi-modal information on the Internet, unimodal search cannot satisfy the requirement of Internet applications. Text-image retrieval research is need…
Is There a One-Model-Fits-All Approach to Information Extraction? Revisiting Task Definition Biases
Wenhao Huang, Qianyu He, Zhixu Li +2
Definition bias is a negative phenomenon that can mislead models. Definition bias in information extraction appears not only across datasets from different domains but also within…
OVEL: Large Language Model as Memory Manager for Online Video Entity Linking
Haiquan Zhao, Xuwu Wang, Shisong Chen +3
In recent years, multi-modal entity linking (MEL) has garnered increasing attention in the research community due to its significance in numerous multi-modal applications. Video, a…
Improving the Robustness of Knowledge-Grounded Dialogue via Contrastive Learning
Jiaan Wang, Jianfeng Qu, Kexin Wang +4
Knowledge-grounded dialogue (KGD) learns to generate an informative response based on a given dialogue context and external knowledge (\emph{e.g.}, knowledge graphs; KGs). Recently…
AspectMMKG: A Multi-modal Knowledge Graph with Aspect-aware Entities
Jingdan Zhang, Jiaan Wang, Xiaodan Wang +2
Multi-modal knowledge graphs (MMKGs) combine different modal data (e.g., text and image) for a comprehensive understanding of entities. Despite the recent progress of large-scale M…
Towards Unifying Multi-Lingual and Cross-Lingual Summarization
Jiaan Wang, Fandong Meng, Duo Zheng +4
To adapt text summarization to the multilingual world, previous work proposes multi-lingual summarization (MLS) and cross-lingual summarization (CLS). However, these two tasks have…