Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
A Tale of LLMs and Induced Small Proxies: Scalable Small Language Models for Knowledge Mining
Sipeng Zhang, Longfei Yun, Shuhuai Lin +8
At the core of Deep Research is knowledge mining, the task of extracting structured information from massive unstructured text in response to user instructions. Large language mode…
cs.AI2026
Relevant Is Not Warranted: Evidence-Force Calibration for Cited RAG
Pin Qian, Su Wang, Xiaoyuan Wang +7
Cited RAG evaluation often treats visible sources as a grounding signal, but a real, topically relevant citation can still under-warrant the attached wording. We study this diagnos…
cs.AI2026
Can RL Teach Long-Horizon Reasoning to LLMs? Expressiveness Is Key
Tianle Wang, Zhaoyang Wang, Guangchen Lan +4
Reinforcement learning (RL) has been applied to improve large language model (LLM) reasoning, yet the systematic study of how training scales with task difficulty has been hampered…