3 papers
cs.CL2026
Route Before Retrieve: Activating Latent Routing Abilities of LLMs for RAG vs. Long-Context Selection
Yiwen Chen, Kuan Li, Fuzhen Zhuang +6
Recent advances in large language models (LLMs) have expanded the context window to beyond 128K tokens, enabling long-document understanding and multi-source reasoning. A key chall…
cs.IR2026
LASAR: Latent Adaptive Semantic Aligned Reasoning for Generative Recommendation
Yiwen Chen, Fuwei Zhang, Zehao Chen +8
Large Language Models (LLMs) have demonstrated powerful reasoning capabilities through Chain-of-Thought (CoT) in various tasks, yet the inefficiency of token-by-token generation hi…
cs.CR2026
Permit: Permission-Aware Representation Intervention for Controlled Generation in Large Language Models
Pengcheng Sun, Lan Zhang, Zhaopeng Zhang +2
Large language models (LLMs) are increasingly deployed in enterprise settings where they handle sensitive documents and user context, raising acute concerns over security and contr…