From the 1 of 4 linked papers with an AI index.
4 papers
Cost-Sensitive Conformal Prediction and Human-in-the-Loop Abstention for Imbalanced High-Stakes Decision Support: A Multi-Domain Benchmark
Manpreet Singh, Akshatha Srikantha, Shyamal Lakhanpal
The paper evaluates how conformal prediction methods can be made cost‑sensitive and reliable for rare, high‑cost classes in imbalanced, high‑stakes decision tasks, and shows that M…
AgentFairBench: Do LLM Agents Discriminate When They Act?
Triveni Morla, Rohith Reddy Bellibaltu, Manpreet Singh +1
Large language model (LLM) agents increasingly take actions (screening applicants, recommending credit, triaging patients), yet fairness for LLMs is still measured by grading answe…
TriEval: A Resource-Efficient Pipeline for LLM Bias, Toxicity, and Truthfulness Assessment
Akshatha Srikantha, Manpreet Singh, Yash Jajoo +1
LLMs have evolved from basic chatbots to the backbone of the AI ecosystem, now widely used in healthcare, schools, and government services. The domain-wide adoption of LLMs necessi…
RISED: A Pre-Deployment Evaluation Framework for High-Stakes AI Decision-Support Systems, with Application to Healthcare
Rohith Reddy Bellibatlu, Manpreet Singh, Yash Jajoo +2
Clinical decision-support systems are expert systems whose recommendations clinicians act on directly, yet they are usually cleared on one aggregate accuracy number from a held-out…