6 papers
Causal state binding predicts action control in language agents
Xiao Jia
Autonomous language agents increasingly expose traces, memories, plans and constraints, but existing evaluations rarely test whether these state variables are bound to final action…
Do Language Models Align with Brains? Prediction Scores Are Not Enough
Xiao Jia
Brain-language model comparisons often interpret neural prediction scores as evidence that model representations capture brain-relevant language computation. We asked whether langu…
NeuroState-Bench: A Human-Calibrated Benchmark for Commitment Integrity in LLM Agent Profiles
Xiao Jia
Outcome-only evaluation under-specifies whether an evaluated agent profile preserves the commitments required to solve a multi-turn task coherently. NeuroState-Bench is a human-cal…
Nautilus: From One Prompt to Plug-and-Play Robot Learning
Yufeng Jin, Jianfei Guo, Xiaogang Jia +8
Robot learning research is fragmented across policy families, benchmark suites, and real robots; each implementation is entangled with the others in a complex combination matrix, m…
The Emergence of Social Science of Large Language Models
Xiao Jia, Zhanzhan Zhao
The social science of large language models (LLMs) examines how these systems evoke mind attributions, interact with one another, and transform human activity and institutions. We…
The Emergence of Altruism in Large-Language-Model Agents Society
Haoyang Li, Xiao Jia, Zhanzhan Zhao
Leveraging Large Language Models (LLMs) for social simulation is a frontier in computational social science. Understanding the social logics these agents embody is critical to this…