Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
STAMP: Provenance-Guided Credit Assignment for Deep Search Agents
Ke Xu, Han Xu, Xinran Chen +6
Reinforcement learning for deep-search agents has largely focused on trajectory-level scoring -- outcome correctness, citation-aware rewards, and evidence coverage. Yet the actions…
cs.AI2025
DoPI: Doctor-like Proactive Interrogation LLM for Traditional Chinese Medicine
Zewen Sun, Ruoxiang Huang, Jiahe Feng +7
Enhancing interrogation capabilities in Traditional Chinese Medicine (TCM) diagnosis through multi-turn dialogues and knowledge graphs presents a significant challenge for modern A…