2 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.MA2024
Role Play: Learning Adaptive Role-Specific Strategies in Multi-Agent Interactions
Weifan Long, Wen Wen, Peng Zhai +1
Zero-shot coordination problem in multi-agent reinforcement learning (MARL), which requires agents to adapt to unseen agents, has attracted increasing attention. Traditional approa…
cs.CL2024★ 2 cited
Phi-3 Safety Post-Training: Aligning Language Models with a "Break-Fix" Cycle
Emman Haider, Daniel Perez-Becker, Thomas Portet +28
Recent innovations in language model training have demonstrated that it is possible to create highly performant models that are small enough to run on a smartphone. As these models…