Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
DSADF: Thinking Fast and Slow for Decision Making
Zhihao Dou, Dongfei Cui, Jun Yan +5
Although Reinforcement Learning (RL) agents are effective in well-defined environments, they often struggle to generalize their learned policies to dynamic settings due to their re…
cs.LG2023
Less or More From Teacher: Exploiting Trilateral Geometry For Knowledge Distillation
Chengming Hu, Haolun Wu, Xuan Li +5
Knowledge distillation aims to train a compact student network using soft supervision from a larger teacher network and hard supervision from ground truths. However, determining an…