Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
ECHO: Entropy-Confidence Hybrid Optimization for Test-Time Reinforcement Learning
Chu Zhao, Enneng Yang, Yuting Liu +2
Test-time reinforcement learning generates multiple candidate answers via repeated rollouts and performs online updates using pseudo-labels constructed by majority voting. To reduc…
cs.LG2025
Graph Representation Learning via Causal Diffusion for Out-of-Distribution Recommendation
Chu Zhao, Enneng Yang, Yuliang Liang +5
Graph Neural Networks (GNNs)-based recommendation algorithms typically assume that training and testing data are drawn from independent and identically distributed (IID) spaces. Ho…