855 citations · 1.2k across the 61 of their papers we have counts for
5 papers · 1 filter
SAW: Stage-Aware Dynamic Weighting for Multi-Objective Reinforcement Learning in Large Language Models
Yuchen He, Baolong Bi, Shenghua Liu +7
Although multi-objective reinforcement learning (MORL) is central to aligning large language models with complex human preferences, the prevailing practice of static weighted summa…
Selective Temporal Knowledge Graph Reasoning
Zhongni Hou, Xiaolong Jin, Zixuan Li +3
Temporal Knowledge Graph (TKG), which characterizes temporally evolving facts in the form of (subject, relation, object, timestamp), has attracted much attention recently. TKG reas…
LegoNet: A Fast and Exact Unlearning Architecture
Sihao Yu, Fei Sun, Jiafeng Guo +2
Machine unlearning aims to erase the impact of specific training samples upon deleted requests from a trained model. Re-training the model on the retained data after deletion is an…
MQGrad: Reinforcement Learning of Gradient Quantization in Parameter Server
Guoxin Cui, Jun Xu, Wei Zeng +3
One of the most significant bottleneck in training large scale machine learning models on parameter server (PS) is the communication overhead, because it needs to frequently exchan…
Locally Smoothed Neural Networks
Liang Pang, Yanyan Lan, Jun Xu +2
Convolutional Neural Networks (CNN) and the locally connected layer are limited in capturing the importance and relations of different local receptive fields, which are often cruci…