attention bias 1dynamic supervision 1out-of-distribution robustness 1robot manipulation 1visual prediction 1
From the 1 of 8 linked papers with an AI index.
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
ReOrder-OPD:Reliability-Aware Prompt Ordering for On-Policy Distillation
Ximo Zhu, Ruiqi Liu, Rong Wang +8
On-policy distillation (OPD) applies token-level teacher supervision to student-generated trajectories, but this supervision is not always reliable. Existing methods use local conf…
cs.LG2026
SGNN: Efficient Global Mixing and Local Message Passing for Long-Range Graph Learning
Dai Shi, Luke Thompson, Linhan Luo +4
Message-passing neural networks (MPNNs) often suffer from an information bottleneck when capturing long-range dependencies, leading to the oversquashing (OSQ) phenomenon. Alongside…