2 papers
cs.LG2026
Beyond Precision: Training-Inference Mismatch is an Optimization Problem and Simple LR Scheduling Fixes It
Yaxiang Zhang, Yingru Li, Jiacai Liu +4
Reinforcement Learning (RL) for training Large Language Models is notoriously unstable. While recent studies attribute this to "training inference mismatch stemming" from inconsist…
cs.SI2025
Applying Machine Learning for characterizing social networks Agent-based models
Haoyuan Li, Lidia Conde Matos, Eduardo César Galobardes +1
Nowadays, social media networks are increasingly significant to our lives, the imperative to study social media networks becomes more and more essential. With billions of users acr…