7 citations · 7 across the 3 of their papers we have counts for
3 papers
cs.CL2025
Smooth Reading: Bridging the Gap of Recurrent LLM to Self-Attention LLM on Long-Context Tasks
Kai Liu, Zhan Su, Peijie Dong +4
Recently, recurrent large language models (Recurrent LLMs) with linear computational complexity have re-emerged as efficient alternatives to self-attention-based LLMs (Self-Attenti…
physics.flu-dyn2024
Comprehensive Study on the Slat Noise of 30P30N High-Lift Airfoil Basd on High-Order Wall-Resolved Large-Eddy Simulation
Keli Zhang, Shizhi Lin, Peiqing Liu +2
This study presents wall-resolved large-eddy simulations (WRLES) of a high-lift airfoil, based on high-order flux reconstruction (FR) commercial software Dimaxer, which runs on con…
cs.CL2024★ 7 cited
Hunyuan-Large: An Open-Source MoE Model with 52 Billion Activated Parameters by Tencent
Xingwu Sun, Yanfeng Chen, Yiqing Huang +105
In this paper, we introduce Hunyuan-Large, which is currently the largest open-source Transformer-based mixture of experts model, with a total of 389 billion parameters and 52 bill…