3 papers
cs.CL2025
Policy-as-Prompt: Turning AI Governance Rules into Guardrails for AI Agents
Gauri Kholkar, Ratinder Ahuja
As autonomous AI agents are used in regulated and safety-critical settings, organizations need effective ways to turn policy into enforceable controls. We introduce a regulatory ma…
cs.CL2025
CAPTURE: Context-Aware Prompt Injection Testing and Robustness Enhancement
Gauri Kholkar, Ratinder Ahuja
Prompt injection remains a major security risk for large language models. However, the efficacy of existing guardrail models in context-aware settings remains underexplored, as the…
cs.CL2024
Socio-Culturally Aware Evaluation Framework for LLM-Based Content Moderation
Shanu Kumar, Gauri Kholkar, Saish Mendke +3
With the growth of social media and large language models, content moderation has become crucial. Many existing datasets lack adequate representation of different groups, resulting…