11 citations · 17 across the 4 of their papers we have counts for
Showing 2023Show all
2 papers · 1 filter
cs.LG2023★ 1 cited
Knowledge Distillation for Efficient Sequences of Training Runs
Xingyu Liu, Alex Leonardi, Lu Yu +3
In many practical scenarios -- like hyperparameter search or continual retraining with new data -- related training runs are performed many times in sequence. Current practice is t…
cs.CV2023★ 2 cited
X-Pruner: eXplainable Pruning for Vision Transformers
Lu Yu, Wei Xiang
Recently vision transformer models have become prominent models for a range of tasks. These models, however, usually suffer from intensive computational costs and heavy memory requ…