2 papers
cs.DC2025
Enhancing Memory Efficiency in Large Language Model Training Through Chronos-aware Pipeline Parallelism
Xinyuan Lin, Chenlu Li, Zongle Huang +5
Larger model sizes and longer sequence lengths have empowered the Large Language Model (LLM) to achieve outstanding performance across various domains. However, this progress bring…
cs.LG2021
Graph Partner Neural Networks for Semi-Supervised Learning on Graphs
Langzhang Liang, Cuiyun Gao, Shiyi Chen +5
Graph Convolutional Networks (GCNs) are powerful for processing graph-structured data and have achieved state-of-the-art performance in several tasks such as node classification, l…