5 papers
Social World Model for Lifelong Social Intelligence
Yu Luo
Social intelligence is a core competency for language agents, yet current research primarily focuses on static capability evaluation rather than how these skills are continuously s…
FLAC: Maximum Entropy RL via Kinetic Energy Regularized Bridge Matching
Lei Lv, Yunfei Li, Yu Luo +2
Iterative generative policies, such as diffusion models and flow matching, offer superior expressivity for continuous control but complicate Maximum Entropy Reinforcement Learning…
Towards Task-Oriented Flying: Framework, Infrastructure, and Principles
Kangyao Huang, Hao Wang, Jingyu Chen +6
Deploying robot learning methods to aerial robots in unstructured environments remains both challenging and promising. While recent advances in deep reinforcement learning (DRL) ha…
Flow-Based Policy for Online Reinforcement Learning
Lei Lv, Yunfei Li, Yu Luo +4
We present \textbf{FlowRL}, a novel framework for online reinforcement learning that integrates flow-based policy representation with Wasserstein-2-regularized optimization. We arg…
Multi-segment Soft Robot Control via Deep Koopman-based Model Predictive Control
Lei Lv, Lei Liu, Lei Bao +7
Soft robots, compared to regular rigid robots, as their multiple segments with soft materials bring flexibility and compliance, have the advantages of safe interaction and dexterou…