Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
CompassDPO: Dynamics-Controlled Direct Preference Optimization for Robust Safety Alignment
Jilong Liu, Yonghui Yang, Pengyang Shao +5
Direct Preference Optimization (DPO) has become a standard framework for safety alignment, but its reliance on pairwise preference updates makes training sensitive to imperfect sup…
cs.LG2025
Distributed Graph Neural Network Inference With Just-In-Time Compilation For Industry-Scale Graphs
Xiabao Wu, Yongchao Liu, Wei Qin +1
Graph neural networks (GNNs) have delivered remarkable results in various fields. However, the rapid increase in the scale of graph data has introduced significant performance bott…