activity
20242026
collaborators

8 papers

cs.AI2026

MonitrLLM: A Community-Centered Evaluation Infrastructure for Large Language Models

Victor Ojewale, Ro Encarnación, Suresh Venkatasubramanian +1

Benchmark suites assess model capability on controlled tasks; large-scale conversation corpora capture naturalistic use without user feedback; and in-interface feedback mechanisms…

cs.CY2026

Bridging Predictions and Interventions: An Integrated Framework for Automated Decision-Systems

Inioluwa Deborah Raji, Lydia T. Liu, Angela Zhou +27

Automated decision systems (ADS) leverage predictions about individual future outcomes to inform consequential decision-making in organizational settings. Across various settings -…

cs.AI2026

Designing for Doubt: The Case for Informed Abstention in Autonomous Agents

Victor Ojewale, Suresh Venkatasubramanian

As large language models gain tool access and are deployed as autonomous agents capable of editing records, executing transactions, and modifying infrastructure, we still evaluate…

cs.CY2026

How to Stop Playing Whack-a-Mole: Mapping the Ecosystem of Technologies Facilitating AI-Generated Non-Consensual Intimate Images

Michelle L. Ding, Harini Suresh, Suresh Venkatasubramanian

The last decade has witnessed a rapid advancement of generative AI technology that significantly scaled the accessibility of AI-generated non-consensual intimate images (AIG-NCII),…

cs.CY2026

Audit Trails for Accountability in Large Language Models

Victor Ojewale, Harini Suresh, Suresh Venkatasubramanian

Large language models (LLMs) are increasingly embedded in consequential decisions across healthcare, finance, employment, and public services. Yet accountability remains fragile be…

cs.LG2025

Bridging Prediction and Intervention Problems in Social Systems

Lydia T. Liu, Inioluwa Deborah Raji, Angela Zhou +32

Many automated decision systems (ADS) are designed to solve prediction problems -- where the goal is to learn patterns from a sample of the population and apply them to individuals…