2 papers
cs.AI2026
Litmus: Zero-Label, Code-Driven Metric Specification for Evaluating AI Systems
Prajjwal Gupta, Prasang Gupta, Vishal Bhutani +4
As agentic LLM systems move from prototypes to deployment across increasingly diverse domains, evaluating them has become both more important and more difficult. The challenge is n…
cs.CR2024
Mean Estimation with User-Level Privacy for Spatio-Temporal IoT Datasets
V. Arvind Rameshwar, Anshoo Tandon, Prajjwal Gupta +3
This paper considers the problem of the private release of sample means of speed values from traffic datasets. Our key contribution is the development of user-level differentially…