#trustworthiness
topictrustworthiness
3 papers · 1 filter
cs.SE2026
ExplainBench: Evaluating Code Explanations from Agents
Zhiyuan Pan, Sungmin Kang, Imam Nur Bani Yusuf +1
The paper introduces ExplainBench, a benchmark that automatically evaluates how trustworthy the explanations generated by code‑writing LLM agents are, by checking if the explanatio…
cs.LG2026★ 8 cited
Towards a Unified Multidimensional Explainability Metric: Evaluating Trustworthiness in AI Models
Georgios Makridis, Georgios Fatouros, Athanasios Kiourtis +4
The paper proposes a framework that evaluates explainability methods like LIME and SHAP across models and datasets using fidelity, simplicity, and stability, and builds a knowledge…
cs.AI2026
Agentic Service-Oriented Computing: A Manifesto for the Next Frontier of Service-Oriented Computing
Amin Beheshti, Rong N. Chang, Boualem Benatallah +7
The paper proposes Agentic Service-Oriented Computing (ASOC), a framework for engineering large‑language‑model‑powered autonomous agents as services and managing their composition,…