Showing 2026Show all
2 papers · 1 filter
cs.AI2026
Measuring and Detecting Harmful AI Sycophancy
Bohan Jiang, Dawei Li, Yasin Silva +1
Sycophantic responses are becoming pervasive in large language models (LLMs), and prior work has pointed out that some of them could be harmful. This paper focuses on one harmful s…
cs.AI2026
ToolPRMBench: Evaluating and Advancing Process Reward Models for Tool-using Agents
Dawei Li, Yuguang Yao, Zhen Tan +2
Reward-guided search methods have demonstrated strong potential in enhancing tool-using agents by effectively guiding sampling and exploration over complex action spaces. As a core…