4 papers
Scaling Up Data Parallelism in Decentralized Deep Learning
Bing Xie, Junqi Yin, Zhenyu Zhou +2
Although it has been extensively explored in theory, decentralized learning is not yet green-lighted for production use, largely due to a lack of stability, scalability, and genera…
Decoding Memories: An Efficient Pipeline for Self-Consistency Hallucination Detection
Weizhi Gao, Xiaorui Liu, Feiyi Wang +2
Large language models (LLMs) have demonstrated impressive performance in both research and real-world applications, but they still struggle with hallucination. Existing hallucinati…
Pixel-Resolved Long-Context Learning for Turbulence at Exascale: Resolving Small-scale Eddies Toward the Viscous Limit
Junqi Yin, Mijanur Palash, M. Paul Laiu +6
Turbulence plays a crucial role in multiphysics applications, including aerodynamics, fusion, and combustion. Accurately capturing turbulence's multiscale characteristics is essent…
Modulated Diffusion: Accelerating Generative Modeling with Modulated Quantization
Weizhi Gao, Zhichao Hou, Junqi Yin +3
Diffusion models have emerged as powerful generative models, but their high computation cost in iterative sampling remains a significant bottleneck. In this work, we present an in-…