3 papers
cs.LG2025
M3-JEPA: Multimodal Alignment via Multi-gate MoE based on the Joint-Embedding Predictive Architecture
Hongyang Lei, Xiaolong Cheng, Qi Qin +8
Current multimodal learning strategies primarily optimize in the original token space. Such a framework is easy to incorporate with the backbone of pretrained language model, but m…
cs.CL2025
LaMsS: When Large Language Models Meet Self-Skepticism
Yetao Wu, Yihong Wang, Teng Chen +4
Hallucination is a major challenge for large language models (LLMs), preventing their further application in some fields. The skeptical thinking of humankind could be useful for LL…
cs.IR2025
Large Language Model Can Be a Foundation for Hidden Rationale-Based Retrieval
Luo Ji, Feixiang Guo, Teng Chen +9
Despite the recent advancement in Retrieval-Augmented Generation (RAG) systems, most retrieval methodologies are often developed for factual retrieval, which assumes query and posi…