Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Continual Distillation of Teachers from Different Domains
Nicolas Michel, Maorong Wang, Jiangpeng He +1
Deep learning models continue to scale, with some requiring more storage than many large-scale datasets. Thus, we introduce a new paradigm: Continual Distillation (CD), where a stu…
cs.LG2025
From Offline to Online Memory-Free and Task-Free Continual Learning via Fine-Grained Hypergradients
Nicolas Michel, Maorong Wang, Jiangpeng He +1
Continual Learning (CL) aims to learn from a non-stationary data stream where the underlying distribution changes over time. While recent advances have produced efficient memory-fr…