Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Budgeted LoRA: Distillation as Structured Compute Allocation for Efficient Inference
Mohammed Sabry, Anya Belz
We study distillation for large language models under explicit compute constraints, with the goal of producing student models that are not only cheaper to train, but structurally e…
cs.LG2024
On the Reduction of Variance and Overestimation of Deep Q-Learning
Mohammed Sabry, Amr M. A. Khalifa
The breakthrough of deep Q-Learning on different types of environments revolutionized the algorithmic design of Reinforcement Learning to introduce more stable and robust algorithm…