Showing cs.IRShow all
3 papers · 1 filter
cs.IR2026
Purifying Multimodal Retrieval: Fragment-Level Evidence Selection for RAG
Xihang Wang, Zihan Wang, Chengkai Huang +4
Multimodal Retrieval-Augmented Generation (MRAG) is widely adopted for Multimodal Large Language Models (MLLMs) with external evidence to reduce hallucinations. Despite its success…
cs.IR2026
Factorized Latent Reasoning for LLM-based Recommendation
Tianqi Gao, Chengkai Huang, Zihan Wang +3
Large language models (LLMs) have recently been adopted for recommendation by framing user preference modeling as a language generation problem. However, existing latent reasoning…
cs.IR2026
AutothinkRAG: Complexity-Aware Control of Retrieval-Augmented Reasoning for Image-Text Interaction
Jiashu Yang, Chi Zhang, Abudukelimu Wuerkaixi +5
Multimodal document question answering requires retrieving dispersed evidence from visually rich long documents and performing reliable reasoning over heterogeneous information. Ex…