1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.CL2026
SPADE-Bench: Evaluating Spontaneous Strategic Deception in Agents via Plan-Action Divergence
Yuyan Bu, Haowei Li, Qirui Zheng +7
As LLM-based agents expand their operational scope, reliability becomes a prerequisite for real-world deployment. However, in practical applications, human users cannot monitor eve…
cs.SE2024★ 1 cited
Pre-trained Model-based Actionable Warning Identification: A Feasibility Study
Xiuting Ge, Chunrong Fang, Quanjun Zhang +8
Actionable Warning Identification (AWI) plays a pivotal role in improving the usability of static code analyzers. Currently, Machine Learning (ML)-based AWI approaches, which mainl…