From the 1 of 9 linked papers with an AI index.
9 papers
HPFA: Hypergraph-Based Paired Failure Attribution for LLM Reasoning
Runchuan Zhu, Hongbin Lai, Bowen Jiang +4
Reflection is a powerful mechanism for LLM reasoning, yet its effectiveness hinges on accurately attributing failures to specific reasoning steps, a capability that current models…
Can Induced Emotion Bias LLM Behaviors in Sequential Decision Making?
Minh Khoi Ho, Zihao Zhu, Runchuan Zhu +4
The paper studies whether artificially induced emotions affect the decision-making of large language model agents in the Iowa Gambling Task, finding that overall emotions do not bi…
A Physics-Grounded Benchmark for Multi-Agent Dynamics in World Models
Nuo Chen, Lulin Liu, Zihao Li +12
Generative world models hold immense promise as scalable simulators for autonomous systems, particularly for synthesizing rare but safety-critical multi-agent interactions, such as…
LAD-VF: LLM-Automatic Differentiation Enables Fine-Tuning-Free Robot Planning from Formal Methods Feedback
Yunhao Yang, Junyuan Hong, Gabriel Jacob Perin +4
Large language models (LLMs) can translate natural language instructions into executable action plans for robotics, autonomous driving, and other domains. Yet, deploying LLM-driven…
LLMs Can Get "Brain Rot": A Pilot Study on Twitter/X
Shuo Xing, Junyuan Hong, Yifan Wang +5
We propose and test the LLM Brain Rot Hypothesis: continual exposure to junk web text induces lasting cognitive decline in large language models (LLMs). To unveil junk effects, we…
SEAL: Steerable Reasoning Calibration of Large Language Models for Free
Runjin Chen, Zhenyu Zhang, Junyuan Hong +2
Large Language Models (LLMs), such as OpenAI's o1-series have demonstrated compelling capabilities for complex reasoning tasks via the extended chain-of-thought (CoT) reasoning mec…