3 papers
cs.LG2025
Adversarial Training for Process Reward Models
Gurusha Juneja, Deepak Nathani, William Yang Wang
Process Reward Models (PRMs) enhance reasoning ability of LLMs by providing step-level supervision. However, their widespread adoption is limited due to expensive manual step-level…
cs.CR2025
MAGPIE: A benchmark for Multi-AGent contextual PrIvacy Evaluation
Gurusha Juneja, Jayanth Naga Sai Pasupulati, Alon Albalak +2
A core challenge for autonomous LLM agents in collaborative settings is balancing robust privacy understanding and preservation alongside task efficacy. Existing privacy benchmarks…
cs.AI2025
MAGPIE: A dataset for Multi-AGent contextual PrIvacy Evaluation
Gurusha Juneja, Alon Albalak, Wenyue Hua +1
The proliferation of LLM-based agents has led to increasing deployment of inter-agent collaboration for tasks like scheduling, negotiation, resource allocation etc. In such systems…