Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
Distilling LLM Reasoning into an Interpretable Policy Tree for Human-AI Collaboration
Beiwen Zhang, Yongheng Liang, Guowei Zou +2
Constructing efficient and reliable policies to assist humans is indispensable for human-AI collaboration. Existing methods mainly follow two lines of work. Most prior work relies…
cs.AI2026
CoFlow: Coordinated Few-Step Flow for Offline Multi-Agent Decision Making
Guowei Zou, Haitao Wang, Beiwen Zhang +2
Generative models have emerged as a promising paradigm for offline multi-agent reinforcement learning (MARL), but existing approaches require many iterative sampling steps. Recent…
cs.AI2025
D2PPO: Diffusion Policy Policy Optimization with Dispersive Loss
Guowei Zou, Weibing Li, Hejun Wu +3
Diffusion policies excel at robotic manipulation by naturally modeling multimodal action distributions in high-dimensional spaces. Nevertheless, diffusion policies suffer from diff…