activity
20242026
collaborators

5 papers

cs.CL2026

Qwen-AgentWorld: Language World Models for General Agents

Yuxin Zuo, Zikai Xiao, Li Sheng +30

A world model predicts environment dynamics based on current observations and actions, serving as a core cognitive mechanism for reasoning and planning. In this work, we investigat…

cs.AI2026

Beyond Safe Data: Pretraining-Stage Alignment with Regular Safety Reflection

Jinhan Li, Kexian Tang, Yihan Xu +2

To achieve deeper safety alignment for large language models (LLMs), recent efforts have studied how to push safety interventions earlier into the pretraining stage, primarily by f…

cs.CL2025

Strategic Planning and Rationalizing on Trees Make LLMs Better Debaters

Danqing Wang, Zhuorui Ye, Xinran Zhao +2

Winning competitive debates requires sophisticated reasoning and argument skills. There are unique challenges in the competitive debate: (1) The time constraints force debaters to…

cs.LG2025

LICORICE: Label-Efficient Concept-Based Interpretable Reinforcement Learning

Zhuorui Ye, Stephanie Milani, Geoffrey J. Gordon +1

Recent advances in reinforcement learning (RL) have predominantly leveraged neural network policies for decision-making, yet these models often lack interpretability, posing challe…

cs.AI2024

Cooperative Strategic Planning Enhances Reasoning Capabilities in Large Language Models

Danqing Wang, Zhuorui Ye, Fei Fang +1

Enhancing the reasoning capabilities of large language models (LLMs) is crucial for enabling them to tackle complex, multi-step problems. Multi-agent frameworks have shown great po…