2 papers
cs.LG2024
When Will Gradient Regularization Be Harmful?
Yang Zhao, Hao Zhang, Xiuyuan Hu
Gradient regularization (GR), which aims to penalize the gradient norm atop the loss function, has shown promising results in training modern over-parameterized deep neural network…
cs.LG2024
Empirical Evidence for the Fragment level Understanding on Drug Molecular Structure of LLMs
Xiuyuan Hu, Guoqing Liu, Yang Zhao +1
AI for drug discovery has been a research hotspot in recent years, and SMILES-based language models has been increasingly applied in drug molecular design. However, no work has exp…