4 papers
TPO: Uncertainty-Guided Exploration Control for Stable Multi-Turn Agentic Reinforcement Learning
Haixin Wang, Hejie Cui, Chenwei Zhang +7
Recent progress in multi-turn reinforcement learning (RL) has significantly improved reasoning LLMs' performances on complex interactive tasks. Despite advances in stabilization te…
LLMs Reading the Rhythms of Daily Life: Aligned Understanding for Behavior Prediction and Generation
Fanjin Meng, Jingtao Ding, Nian Li +2
Human daily behavior unfolds as complex sequences shaped by intentions, preferences, and context. Effectively modeling these behaviors is crucial for intelligent systems such as pe…
CORGI: GNNs with Convolutional Residual Global Interactions for Lagrangian Simulation
Ethan Ji, Yuanzhou Chen, Arush Ramteke +5
Partial differential equations (PDEs) are central to dynamical systems modeling, particularly in hydrodynamics, where traditional solvers often struggle with nonlinearity and compu…
Why Are We Moral? An LLM-based Agent Simulation Approach to Study Moral Evolution
Zhou Ziheng, Huacong Tang, Mingjie Bi +7
The evolution of morality presents a puzzle: natural selection should favor self-interest, yet humans developed moral systems promoting altruism. Traditional approaches must abstra…