3 papers
cs.LG2026
More Expressive Feedforward Layers: Part I. Token-Adaptive Mixing of Activations
Mingze Wang, Jinbo Wang, Yikuan Xia +2
Feedforward network (FFN) layers account for a large fraction of parameters and nonlinear expressivity in Transformer-based large language models (LLMs). Despite the evolution from…
cs.IR2026
DB3 Team's Solution For Meta KDD Cup' 25
Yikuan Xia, Jiazun Chen, Yirui Zhan +6
This paper presents the db3 team's winning solution for the Meta CRAG-MM Challenge 2025 at KDD Cup'25. Addressing the challenge's unique multi-modal, multi-turn question answering…
cs.IR2025
ER-RAG: Enhance RAG with ER-Based Unified Modeling of Heterogeneous Data Sources
Yikuan Xia, Jiazun Chen, Yirui Zhan +6
Large language models (LLMs) excel in question-answering (QA) tasks, and retrieval-augmented generation (RAG) enhances their precision by incorporating external evidence from diver…