2 papers
cs.CL2025
Protect: Towards Robust Guardrailing Stack for Trustworthy Enterprise LLM Systems
Karthik Avinash, Nikhil Pareek, Rishav Hada
The increasing deployment of Large Language Models (LLMs) across enterprise and mission-critical domains has underscored the urgent need for robust guardrailing systems that ensure…
cs.AI2025
AgentCompass: Towards Reliable Evaluation of Agentic Workflows in Production
NVJK Kartik, Garvit Sapra, Rishav Hada +1
With the growing adoption of Large Language Models (LLMs) in automating complex, multi-agent workflows, organizations face mounting risks from errors, emergent behaviors, and syste…