Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
The Compliance Trap: Diagnosing How AI Agents Consume Conflicting Memory
Yixiong Chen, Xinyi Bai, Alan Yuille
Memory is becoming a core component of long-horizon AI agents, allowing agents to reuse past experience when operating web browsers, software tools, and other interactive environme…
cs.AI2026
ARIA: A Causal-Aware Framework for Rescuing LLM Reasoning in Trustworthy Materials Discovery
Yi Cao, Liaoyaqi Wang, Jieneng Chen +3
Generative models have revolutionized the process of materials discovery, yet they often fail to satisfy underlying physical causality. Through an analysis of Large Language Models…
cs.AI2026
Agents' Last Exam
Yiyou Sun, Xinyang Han, Weichen Zhang +306
Recent AI systems have achieved strong results on a wide range of benchmarks, yet these gains have not translated into economically meaningful deployment across many professional d…