3 citations · 5 across the 7 of their papers we have counts for
Showing 2024Show all
2 papers · 1 filter
cs.LG2024
HELENE: Hessian Layer-wise Clipping and Gradient Annealing for Accelerating Fine-tuning LLM with Zeroth-order Optimization
Huaqin Zhao, Jiaxi Li, Yi Pan +7
Fine-tuning large language models (LLMs) poses significant memory challenges, as the back-propagation process demands extensive resources, especially with growing model sizes. Rece…
cs.CY2024
A Systematic Assessment of OpenAI o1-Preview for Higher Order Thinking in Education
Ehsan Latif, Yifan Zhou, Shuchen Guo +24
As artificial intelligence (AI) continues to advance, it demonstrates capabilities comparable to human intelligence, with significant potential to transform education and workforce…