Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Distilling LLM Reasoning into an Interpretable Policy Tree for Human-AI Collaboration
Beiwen Zhang, Yongheng Liang, Guowei Zou +2
Constructing efficient and reliable policies to assist humans is indispensable for human-AI collaboration. Existing methods mainly follow two lines of work. Most prior work relies…
cs.AI2026
CoFlow: Coordinated Few-Step Flow for Offline Multi-Agent Decision Making
Guowei Zou, Haitao Wang, Beiwen Zhang +2
Generative models have emerged as a promising paradigm for offline multi-agent reinforcement learning (MARL), but existing approaches require many iterative sampling steps. Recent…