2 papers
cs.CL2023
Sensi-BERT: Towards Sensitivity Driven Fine-Tuning for Parameter-Efficient BERT
Souvik Kundu, Sharath Nittur Sridhar, Maciej Szankin +1
Large pre-trained language models have recently gained significant traction due to their improved performance on various down-stream tasks like text classification and question ans…
cs.LG2023
InstaTune: Instantaneous Neural Architecture Search During Fine-Tuning
Sharath Nittur Sridhar, Souvik Kundu, Sairam Sundaresan +2
One-Shot Neural Architecture Search (NAS) algorithms often rely on training a hardware agnostic super-network for a domain specific task. Optimal sub-networks are then extracted fr…