2 papers
cs.IR2026
Rethinking Agentic RAG: Toward LLM-Driven Logical Retrieval Beyond Embeddings
Yuqi Zeng, Qixiang Deng, Yulei Wan +3
Recent advances in RAG have shifted toward an agentic paradigm, where LLMs interact with retrieval systems over multiple turns and iteratively refine queries based on intermediate…
cs.LG2026
Reward Hacking in the Era of Large Models: Mechanisms, Emergent Misalignment, Challenges
Xiaohua Wang, Muzhao Tian, Yuqi Zeng +20
Reinforcement Learning from Human Feedback (RLHF) and related alignment paradigms have become central to steering large language models (LLMs) and multimodal large language models…