3 citations · 3 across the 14 of their papers we have counts for
4 papers · 1 filter
Memory Shot for Long-Term Dialogue
Chunyi Peng, Haidong Xin, Xuanshuo Sheng +7
Large Language Models (LLMs) have demonstrated strong capabilities in general conversation, instruction following, and complex reasoning. However, in long-term dialogue settings, t…
ReAlign: Optimizing the Visual Document Retriever with Reasoning-Guided Fine-Grained Alignment
Hao Yang, Yifan Ji, Zhipeng Xu +6
Visual document retrieval aims to retrieve a set of document pages relevant to a query from visually rich collections. Existing methods often employ Vision-Language Models (VLMs) t…
LISRec: Modeling User Preferences with Learned Item Shortcuts for Sequential Recommendation
Haidong Xin, Zhenghao Liu, Sen Mei +7
User-item interaction histories are pivotal for sequential recommendation systems but often include noise, such as unintended clicks or actions that fail to reflect genuine user pr…
VisRAG: Vision-based Retrieval-augmented Generation on Multi-modality Documents
Shi Yu, Chaoyue Tang, Bokai Xu +8
Retrieval-augmented generation (RAG) is an effective technique that enables large language models (LLMs) to utilize external knowledge sources for generation. However, current RAG…