2 papers
cs.CR2025
Enterprise AI Must Enforce Participant-Aware Access Control
Shashank Shreedhar Bhatt, Tanmay Rajore, Khushboo Aggarwal +10
Large language models (LLMs) are increasingly deployed in enterprise settings where they interact with multiple users and are trained or fine-tuned on sensitive internal data. Whil…
cs.CR2024
TRUCE: Private Benchmarking to Prevent Contamination and Improve Comparative Evaluation of LLMs
Tanmay Rajore, Nishanth Chandran, Sunayana Sitaram +4
Benchmarking is the de-facto standard for evaluating LLMs, due to its speed, replicability and low cost. However, recent work has pointed out that the majority of the open source b…