Showing cs.AIShow all
2 papers · 1 filter
cs.AI2025
NeurIPS 2025 E2LM Competition : Early Training Evaluation of Language Models
Mouadh Yagoubi, Yasser Dahou, Billel Mokeddem +12
Existing benchmarks have proven effective for assessing the performance of fully trained large language models. However, we find striking differences in the early training stages o…
cs.AI2023
Regularization of the policy updates for stabilizing Mean Field Games
Talal Algumaei, Ruben Solozabal, Reda Alami +3
This work studies non-cooperative Multi-Agent Reinforcement Learning (MARL) where multiple agents interact in the same environment and whose goal is to maximize the individual retu…