collaborators

5 papers

cs.CL2026

Counterfactual Fairness Audits of Multi-Step Clinical LLM Agents Require a Measured Per-Action Instability Floor

Rohith Reddy Bellibatlu, Manpreet Singh, Deepak Parashar +1

Counterfactual audits are the standard tool for checking whether a clinical agent treats demographically distinct but clinically identical patients differently. They report a flip…

cs.LG2026

Cost-Sensitive Conformal Prediction and Human-in-the-Loop Abstention for Imbalanced High-Stakes Decision Support: A Multi-Domain Benchmark

Manpreet Singh, Akshatha Srikantha, Shyamal Lakhanpal

High-stakes decision systems in credit scoring, fraud detection, healthcare, and industrial safety require reliable uncertainty quantification under severe class imbalance and asym…

cs.AI2026

TriEval: A Resource-Efficient Pipeline for LLM Bias, Toxicity, and Truthfulness Assessment

Akshatha Srikantha, Manpreet Singh, Yash Jajoo +1

LLMs have evolved from basic chatbots to the backbone of the AI ecosystem, now widely used in healthcare, schools, and government services. The domain-wide adoption of LLMs necessi…

cs.AI2026

AgentFairBench: Do LLM Agents Discriminate When They Act?

Triveni Morla, Rohith Reddy Bellibatlu, Manpreet Singh +1

Large language model (LLM) agents increasingly take actions (screening applicants, recommending credit, triaging patients), yet fairness for LLMs is still measured by grading answe…

cs.LG2026

RISED: A Pre-Deployment Evaluation Framework for High-Stakes AI Decision-Support Systems, with Application to Healthcare

Rohith Reddy Bellibatlu, Manpreet Singh, Yash Jajoo +2

Clinical decision-support systems are expert systems whose recommendations clinicians act on directly, yet they are usually cleared on one aggregate accuracy number from a held-out…