Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
WEQA: Wearable hEalth Question Answering with Query-Adaptive Agentic Reasoning
Yuwei Zhang, Tong Xia, Bianca Emmerich +5
Language models are remarkably capable at medical question answering, in some cases surpassing the accuracy of general physicians. However, answering questions about wearable healt…
cs.AI2026
Controllable and Verifiable Tool-Use Data Synthesis for Agentic Reinforcement Learning
Siyuan Xu, Shiyang Li, Xin Liu +9
Existing synthetic tool-use corpora are primarily designed for offline supervised fine-tuning, yet reinforcement learning (RL) requires executable environments that support reward-…