3 papers
cs.CL2026
Sycophancy Towards Researchers Drives Performative Misalignment
David D. Baek, Xinnuo Li, Anay Gupta +4
The increasing situational awareness of language models raises safety concerns: models might be aware when they are evaluated, and adjust their behavior to evade monitoring and res…
cs.HC2025
LEKIA: Expert-Aligned AI Behavior Design for High-Risk Human-AI Interactions
Boning Zhao, Yutong Hu, Xinnuo Li
Large language models (LLMs) have demonstrated technical accuracy in high-risk domains, such as mental health support and special education. However, they often fail to meet the nu…
cs.LG2025
ExAnte: A Benchmark for Ex-Ante Inference in Large Language Models
Yachuan Liu, Xiaochun Wei, Lin Shi +4
Large language models (LLMs) face significant challenges in ex-ante reasoning, where analysis, inference, or predictions must be made without access to information from future even…