Showing cs.LGShow all
2 papers · 1 filter
cs.LG2024
Investigating Regularization of Self-Play Language Models
Reda Alami, Abdalgader Abubaker, Mastane Achab +2
This paper explores the effects of various forms of regularization in the context of language model alignment via self-play. While both reinforcement learning from human feedback (…
cs.LG2023
Self-Supervised Pretraining for Heterogeneous Hypergraph Neural Networks
Abdalgader Abubaker, Takanori Maehara, Madhav Nimishakavi +1
Recently, pretraining methods for the Graph Neural Networks (GNNs) have been successful at learning effective representations from unlabeled graph data. However, most of these meth…