Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
Wnuan: Staged Post-Training for Question Answering over Proprietary Enterprise Knowledge
Xiaofeng Shi, Xiaosong Qiu, Wenxin Ma +6
Enterprise question answering requires models to acquire proprietary knowledge without discarding general capabilities. We present Wnuan, a three-stage pipeline that constructs tas…
cs.AI2026
Closing the Feedback Loop: From Experience Extraction to Insight Governance in Verbal Reinforcement Learning
Yanwei Cui, Xing Zhang, Yulong Zhang +4
Training-free verbal reinforcement learning enables LLM agents to learn from world feedback -- objective signals such as dynamic task outcomes, market returns, or demand forecasts…
cs.AI2025
SciSage: A Multi-Agent Framework for High-Quality Scientific Survey Generation
Xiaofeng Shi, Qian Kou, Yuduo Li +5
The rapid growth of scientific literature demands robust tools for automated survey-generation. However, current large language model (LLM)-based methods often lack in-depth analys…