Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
CausalGame: Benchmarking Causal Thinking of LLM Agents in Games
Zhenhao Chen, Yongqiang Chen, Chenxi Liu +7
Building AI Scientist agents with Large Language Models (LLMs) has recently attracted growing attention. Since scientific discovery fundamentally relies on uncovering causal relati…
cs.CL2025
On the Thinking-Language Modeling Gap in Large Language Models
Chenxi Liu, Yongqiang Chen, Tongliang Liu +3
System 2 reasoning is one of the defining characteristics of intelligence, which requires slow and logical thinking. Human conducts System 2 reasoning via the language of thoughts…