2 citations · 2 across the 3 of their papers we have counts for
3 papers
cs.DC2025
MegatronApp: Efficient and Comprehensive Management on Distributed LLM Training
Bohan Zhao, Guang Yang, Shuo Chen +4
The rapid escalation in the parameter count of large language models (LLMs) has transformed model training from a single-node endeavor into a highly intricate, cross-node activity.…
hep-th2025★ 2 cited
Island rules for the noncommutative black hole
Yipeng Liu, Wei Xu, Baocheng Zhang
In the context of noncommutative black holes, we reconsider the island rule and reproduce the Page curve. Since the radiation entropy of the noncommutative black hole will eventual…
cs.CL2024
WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Zehan Qi, Xiao Liu, Iat Long Iong +11
Large language models (LLMs) have shown remarkable potential as autonomous agents, particularly in web-based tasks. However, existing LLM web agents heavily rely on expensive propr…