2 papers
cs.AI2026
RAIL Guard: Closing the Evaluation-to-Remediation Gap in Responsible AI for LLM Agents
Sumit Verma, Pritam Prasun, Pritish Kumar
Existing guardrail systems for large language model agents operate as binary classifiers that block unsafe content, leaving organizations to discard failing outputs and retry from…
cs.AI2025
RAIL in the Wild: Operationalizing Responsible AI Evaluation Using Anthropic's Value Dataset
Sumit Verma, Pritam Prasun, Arpit Jaiswal +1
As AI systems become embedded in real-world applications, ensuring they meet ethical standards is crucial. While existing AI ethics frameworks emphasize fairness, transparency, and…