2 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.LG2024
APOLLO: SGD-like Memory, AdamW-level Performance
Hanqing Zhu, Zhenyu Zhang, Wenyan Cong +7
Large language models (LLMs) are notoriously memory-intensive during training, particularly with the popular AdamW optimizer. This memory burden necessitates using more or higher-e…
cs.LG2024★ 2 cited
PACE: Pacing Operator Learning to Accurate Optical Field Simulation for Complicated Photonic Devices
Hanqing Zhu, Wenyan Cong, Guojin Chen +4
Electromagnetic field simulation is central to designing, optimizing, and validating photonic devices and circuits. However, costly computation associated with numerical simulation…