3 papers
cs.LG2026
OGLS-SD: On-Policy Self-Distillation with Outcome-Guided Logit Steering for LLM Reasoning
Yuxiao Yang, Xiaoyun Wang, Weitong Zhang
We study on-policy self-distillation (OPSD), where a language model improves its reasoning ability by distilling privileged teacher distributions along its own on-policy trajectori…
cs.CV2026
SARe: Structure-Aware Generative 3D Fragment Reassembly
Hanze Jia, Chunshi Wang, Yuxiao Yang +4
3D fragment reassembly estimates the rigid pose of each fragment to recover a complete object from unordered point clouds or meshes. The task becomes increasingly challenging as th…
cs.LG2025
Neurocircuitry-Inspired Hierarchical Graph Causal Attention Networks for Explainable Depression Identification
Weidao Chen, Yuxiao Yang, Yueming Wang
Major Depressive Disorder (MDD), affecting millions worldwide, exhibits complex pathophysiology manifested through disrupted brain network dynamics. Although graph neural networks…