Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Faster Verified Explanations for Neural Networks
Alessandro De Palma, Greta Dolcetti, Caterina Urban
Verified explanations are a principled way to explain the decisions taken by neural networks, which are otherwise black-box in nature. However, these techniques face significant sc…
cs.LG2025
Building a Foundational Guardrail for General Agentic Systems via Synthetic Data
Yue Huang, Hang Hua, Yujun Zhou +11
While LLM agents can plan multi-step tasks, intervening at the planning stage-before any action is executed-is often the safest way to prevent harm, since certain risks can lead to…