4 papers
Beyond the Answer: Decoding the Behavior of LLMs as Scientific Reasoners
Rohan Pandey, Eric Ye, Michael Li
As Large Language Models (LLMs) achieve increasingly sophisticated performance on complex reasoning tasks, current architectures serve as critical proxies for the internal heuristi…
Quantization Blindspots: How Model Compression Breaks Backdoor Defenses
Rohan Pandey, Eric Ye
Backdoor attacks embed input-dependent malicious behavior into neural networks while preserving high clean accuracy, making them a persistent threat for deployed ML systems. At the…
Modeling Others' Minds as Code
Kunal Jha, Aydan Yuenan Huang, Eric Ye +2
Accurate prediction of human behavior is essential for robust and safe human-AI collaboration. However, existing approaches for modeling people are often data-hungry and brittle be…
An Efficient Open World Environment for Multi-Agent Social Learning
Eric Ye, Ren Tao, Natasha Jaques
Many challenges remain before AI agents can be deployed in real-world environments. However, one virtue of such environments is that they are inherently multi-agent and contain hum…