1 citations · 1 across the 7 of their papers we have counts for
12 papers
Scalable High-Fidelity Macromolecular Docking for GPU-Accelerated Supercomputers
Xiangyu Meng, Peng Chen, Mingzhen Li +7
Flexible macromolecular docking offers high-fidelity predictions of biomolecular interactions, but remains prohibitively expensive at scale. Among existing approaches, LightDock le…
Direct Model State Migration for Elastic Training of Large Language Models
Weijian Liu, Mingzhen Li, Rui Kang +3
Large language model (LLM) training shall adapt to dynamic resources in shared clusters to tackle the elasticity, including passive preemption and optimistic scaling. State migrati…
Deep Learning-Enabled Supercritical Flame Simulation at Detailed Chemistry and Real-Fluid Accuracy Towards Trillion-Cell Scale
Zhuoqiang Guo, Runze Mao, Lijun Liu +3
For decades, supercritical flame simulations incorporating detailed chemistry and real-fluid transport have been limited to millions of cells, constraining the resolved spatial and…
WISCA: A Lightweight Model Transition Method to Improve LLM Training via Weight Scaling
Jiacheng Li, Jianchao Tan, Zhidong Yang +11
Transformer architecture gradually dominates the LLM field. Recent advances in training optimization for Transformer-based large language models (LLMs) primarily focus on architect…
Large-scale Neural Network Quantum States for ab initio Quantum Chemistry Simulations on Fugaku
Hongtao Xu, Zibo Wu, Mingzhen Li +1
Solving quantum many-body problems is one of the fundamental challenges in quantum chemistry. While neural network quantum states (NQS) have emerged as a promising computational to…
Skrull: Towards Efficient Long Context Fine-tuning through Dynamic Data Scheduling
Hongtao Xu, Wenting Shen, Yuanxin Wei +6
Long-context supervised fine-tuning (Long-SFT) plays a vital role in enhancing the performance of large language models (LLMs) on long-context tasks. To smoothly adapt LLMs to long…