activity
20242026
collaborators

5 papers

cs.AI2026

UniMoMo: Expert Merging-Based MoE Acceleration for Large Recommendation Models

Lei Xin, Bin Gu, Peize Li +9

Sparse mixture-of-experts (MoE) layers expand recommendation capacity through conditional computation, yet a trained checkpoint still stores and routes over its full expert bank. W…

cs.LG2026

Token Reduction Should Go Beyond Efficiency in Generative Models -- From Vision, Language to Multimodality

Zhenglun Kong, Yize Li, Fanhu Zeng +7

In Transformer architectures, tokens\textemdash discrete units derived from raw data\textemdash are formed by segmenting inputs into fixed-length chunks. Each token is then mapped…

cs.IR2026

HyTRec: A Hybrid Temporal-Aware Attention Architecture for Long Behavior Sequential Recommendation

Lei Xin, Yuhao Zheng, Ke Cheng +3

Modeling long sequences of user behaviors has emerged as a critical frontier in generative recommendation. However, existing solutions face a dilemma: linear attention mechanisms a…

cs.LG2025

Life-Code: Central Dogma Modeling with Multi-Omics Sequence Unification

Zicheng Liu, Siyuan Li, Zhiyuan Chen +6

The interactions between DNA, RNA, and proteins are fundamental to biological processes, as illustrated by the central dogma of molecular biology. Although modern biological pre-tr…

q-bio.GN2024

Artificial Intelligence for Central Dogma-Centric Multi-Omics: Challenges and Breakthroughs

Lei Xin, Caiyun Huang, Hao Li +8

With the rapid development of high-throughput sequencing platforms, an increasing number of omics technologies, such as genomics, metabolomics, and transcriptomics, are being appli…