2 papers
cs.LG2025
An Extra RMSNorm is All You Need for Fine Tuning to 1.58 Bits
Cody Steinmetz, Gavin Childress, Aaron Herbst +4
Large language models (LLMs) have transformed natural-language processing, yet their scale makes real-world deployment costly. Post-training quantization reduces memory and computa…
math.NA2024
Gradient Preserving Operator Inference: Data-Driven Reduced-Order Models for Equations with Gradient Structure
Yuwei Geng, Jasdeep Singh, Lili Ju +2
Hamiltonian Operator Inference has been introduced in [Sharma, H., Wang, Z., Kramer, B., Physica D: Nonlinear Phenomena, 431, p.133122, 2022] to learn structure-preserving reduced-…