activity
20242026
collaborators

6 papers

cs.LG2026

RRISE: Robust Radius Inference via a Surrogate Estimator

Jong-Ik Park, Shreyas Chaudhari, Carlee Joe-Wong +1

Randomized smoothing (RS) uses a smoothed classifier to provide architecture-agnostic certificates of classification robustness, but its dependence on per-input Monte Carl…

cs.LG2025

Which Rewards Matter? Reward Selection for Reinforcement Learning under Limited Feedback

Shreyas Chaudhari, Renhao Zhang, Philip S. Thomas +1

The ability of reinforcement learning algorithms to learn effective policies is determined by the rewards available during training. However, for practical problems, obtaining larg…

cs.CL2025

PersonaGym: Evaluating Persona Agents and LLMs

Vinay Samuel, Henry Peng Zou, Yue Zhou +6

Persona agents, which are LLM agents conditioned to act according to an assigned persona, enable contextually rich and user aligned interactions across domains like education and h…

cs.CL2025

Probing AI Safety with Source Code

Ujwal Narayan, Shreyas Chaudhari, Ashwin Kalyan +4

Large language models (LLMs) have become ubiquitous, interfacing with humans in numerous safety-critical applications. This necessitates improving capabilities, but importantly cou…

cs.AI2025

Agent Context Protocols Enhance Collective Inference

Devansh Bhardwaj, Arjun Beniwal, Shreyas Chaudhari +5

AI agents have become increasingly adept at complex tasks such as coding, reasoning, and multimodal understanding. However, building generalist systems requires moving beyond indiv…

cs.LG2024

Abstract Reward Processes: Leveraging State Abstraction for Consistent Off-Policy Evaluation

Shreyas Chaudhari, Ameet Deshpande, Bruno Castro da Silva +1

Evaluating policies using off-policy data is crucial for applying reinforcement learning to real-world problems such as healthcare and autonomous driving. Previous methods for off-…