4 papers
HyReC: Exploring Hybrid-based Retriever for Chinese
Zunran Wang, Zheng Shenpeng, Wang Shenglan +2
Hybrid-based retrieval methods, which unify dense-vector and lexicon-based retrieval, have garnered considerable attention in the industry due to performance enhancement. However,…
Detecting the Root Cause Code Lines in Bug-Fixing Commits by Heterogeneous Graph Learning
Liguo Ji, Chenchen Li, Shenglin Wang +1
With the continuous growth in the scale and complexity of software systems, defect remediation has become increasingly difficult and costly. Automated defect prediction tools can p…
MFC-Bench: Benchmarking Multimodal Fact-Checking with Large Vision-Language Models
Shengkang Wang, Hongzhan Lin, Ziyang Luo +3
Large vision-language models (LVLMs) have significantly improved multimodal reasoning tasks, such as visual question answering and image captioning. These models embed multimodal f…
GT2Vec: Large Language Models as Multi-Modal Encoders for Text and Graph-Structured Data
Jiacheng Lin, Kun Qian, Haoyu Han +9
Graph-structured information offers rich contextual information that can enhance language models by providing structured relationships and hierarchies, leading to more expressive e…