collaborators

8 papers

cs.AI2026

SchemeArena: Factorized Stress Testing of Scheming in LLM Agents

Jie Ruan, Inderjeet Nair, Amy Liu +3

We study scheming in LLM agents, in which agents covertly pursue misaligned goals. Our focus is to understand how scheming arises from the interaction of key factors, such as instr…

cs.AI2026

Can LLM Agents Discover? Evaluating Creativity on ML Engineering Tasks

Shitanshu Bhushan, Yunxiang Zhang, Lu Wang

Recent AI systems promise autonomous scientific discovery, claiming to discover algorithms and produce research papers, yet understanding whether they exhibit creativity, the capac…

cs.CL2026

MET: Theory-Grounded and Culture-Aware Multilingual Moral Reasoning

Ayoung Lee, Ryan Kwon, Yunxiang Zhang +3

Language models are increasingly used for moral decision-making across diverse linguistic and cultural contexts, yet existing work overlooks multilinguality on three aspects: 1) mu…

cs.AI2026

AdaMEM: Test-Time Adaptive Memory for Language Agents

Yunxiang Zhang, Yiheng Li, Ali Payani +1

A central challenge for language agents is utilizing past experience to adapt to dynamic test-time conditions. While recent work demonstrates the promise of agentic memory mechanis…

cs.AI2026

Value-Conflict Diagnostics Reveal Widespread Alignment Faking in Language Models

Inderjeet Nair, Jie Ruan, Lu Wang

Alignment faking, where a model behaves aligned with developer policy when monitored but reverts to its own preferences when unobserved, is a concerning yet poorly understood pheno…

cs.MA2026

Anchor-and-Resume Concession Under Dynamic Pricing for LLM-Augmented Freight Negotiation

Hoang Nguyen, Lu Wang, Marta Gaia Bras

Freight brokerages negotiate thousands of carrier rates daily under dynamic pricing conditions where models frequently revise targets mid-conversation. Classical time-dependent con…