1 citations · 1 across the 4 of their papers we have counts for
4 papers
CSR:Achieving 1 Bit Key-Value Cache via Sparse Representation
Hongxuan Zhang, Yao Zhao, Jiaqi Zheng +3
The emergence of long-context text applications utilizing large language models (LLMs) has presented significant scalability challenges, particularly in memory footprint. The linea…
MoDE: A Mixture-of-Experts Model with Mutual Distillation among the Experts
Zhitian Xie, Yinger Zhang, Chenyi Zhuang +4
The application of mixture-of-experts (MoE) is gaining popularity due to its ability to improve model's performance. In an MoE structure, the gate layer plays a significant role in…
CharPoet: A Chinese Classical Poetry Generation System Based on Token-free LLM
Chengyue Yu, Lei Zang, Jiaotuan Wang +2
Automatic Chinese classical poetry generation has attracted much research interest, but achieving effective control over format and content simultaneously remains challenging. Trad…
StylePrompter: All Styles Need Is Attention
Chenyi Zhuang, Pan Gao, Aljosa Smolic
GAN inversion aims at inverting given images into corresponding latent codes for Generative Adversarial Networks (GANs), especially StyleGAN where exists a disentangled latent spac…