2 papers
cs.SE2026
The "I Don't Know" Filter: Enhancing Agentic Reliability in Function Calling
Stefan Broecker, Mason del Rosario, Boris Selitser +1
The language models that underpin agents have seen a rapid rise in performance on function calling benchmarks. However, the metrics used in the training and evaluation of these mod…
cs.AI2026
Forgetting-MarI: LLM Unlearning via Marginal Information Regularization
Shizhou Xu, Yuan Ni, Stefan Broecker +1
As AI models are trained on ever-expanding datasets, the ability to remove the influence of specific data from trained models has become essential for privacy protection and regula…