1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.LG2026
Pretraining Curricula Enable Selective Fine-tuning
Sebastian A. Bruijns, Jirko Rubruck, Mia H. Whitefield +3
Transformers follow implicit curricula whereby some tasks are learned before others. However, how explicit pretraining curricula influence learning, generalization, and the selecti…
cs.LG2024★ 1 cited
Early learning of the optimal constant solution in neural networks and humans
Jirko Rubruck, Jan P. Bauer, Andrew Saxe +1
Deep neural networks learn increasingly complex functions over the course of training. Here, we show both empirically and theoretically that learning of the target function is prec…