2 papers
cs.LG2026
TENP: Trapezoidal Expert Neuron Pruning For Mixture-of-Experts
Jiangyang He, Shaolin Zhu, Deyi Xiong
Mixture-of-Experts large language models (LLMs) scale efficiently through sparse activation, yet their deployment is fundamentally constrained by the large static parameter footpri…
cs.CL2025
Cultivating Game Sense for Yourself: Making VLMs Gaming Experts
Wenxuan Lu, Jiangyang He, Zhanqiu Zhang +2
Developing agents capable of fluid gameplay in first/third-person games without API access remains a critical challenge in Artificial General Intelligence (AGI). Recent efforts lev…