2 papers
cs.LG2025
Label Smoothing Improves Gradient Ascent in LLM Unlearning
Zirui Pang, Hao Zheng, Zhijie Deng +3
LLM unlearning has emerged as a promising approach, aiming to enable models to forget hazardous/undesired knowledge at low cost while preserving as much model utility as possible.…
cs.MA2025
Network Topology and Information Efficiency of Multi-Agent Systems: Study based on MARL
Xinren Zhang, Sixi Cheng, Zixin Zhong +1
Multi-agent systems (MAS) solve complex problems through coordinated autonomous entities with individual decision-making capabilities. While Multi-Agent Reinforcement Learning (MAR…