2 papers
cs.CR2026
ADVERSA: Measuring Multi-Turn Guardrail Degradation and Judge Reliability in Large Language Models
Harry Owiredu-Ashley
Most adversarial evaluations of large language model (LLM) safety assess single prompts and report binary pass/fail outcomes, which fails to capture how safety properties evolve un…
cs.CY2023
Securing Bystander Privacy in Mixed Reality While Protecting the User Experience
Matthew Corbett, Brendan David-John, Jiacheng Shang +2
The modern Mixed Reality devices that make the Metaverse viable require vast information about the physical world and can also violate the privacy of unsuspecting or unwilling byst…