co-evolution training 1computer-use agents 1reinforcement learning 1stateful applications 1synthetic environments 1
From the 1 of 8 linked papers with an AI index.
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Scaling Reasoning Efficiently via Relaxed On-Policy Distillation
Jongwoo Ko, Sara Abdali, Young Jin Kim +2
On-policy distillation is pivotal for transferring reasoning capabilities to capacity-constrained models, yet remains prone to instability and negative transfer. We show that on-po…
cs.LG2025
Hierarchical Self-Attention: Generalizing Neural Attention Mechanics to Multi-Scale Problems
Saeed Amizadeh, Sara Abdali, Yinheng Li +1
Transformers and their attention mechanism have been revolutionary in the field of Machine Learning. While originally proposed for the language data, they quickly found their way t…