4 papers
Multi-Agent Coordination Adaptation via Structure-Guided Orchestration
Haoran Li, Shulun Chen, Shaoyuan Sun +1
As large language model (LLM)-based multi-agent systems scale to handle increasingly complex tasks, balancing structural stability and dynamic adaptability becomes increasingly cha…
PRISM: A Geometric Risk Bound that Decomposes Drift into Scale, Shape, and Head
Chieh-Yen Lin, Shao-Hua Sun
Comparing post-training LLM variants, such as quantized, LoRA-adapted, and distilled models, requires a diagnostic that identifies how a variant has drifted, not only whether it ha…
Restoring Noisy Demonstration for Imitation Learning With Diffusion Models
Shang-Fu Chen, Co Yong, Shao-Hua Sun
Imitation learning (IL) aims to learn a policy from expert demonstrations and has been applied to various applications. By learning from the expert policy, IL methods do not requir…
HERO: Human-Feedback Efficient Reinforcement Learning for Online Diffusion Model Finetuning
Ayano Hiranaka, Shang-Fu Chen, Chieh-Hsin Lai +6
Controllable generation through Stable Diffusion (SD) fine-tuning aims to improve fidelity, safety, and alignment with human guidance. Existing reinforcement learning from human fe…