2 papers
cs.CV2025
Connecting Domains and Contrasting Samples: A Ladder for Domain Generalization
Tianxin Wei, Yifan Chen, Xinrui He +2
Distribution shifts between training and testing samples frequently occur in practice and impede model generalization performance. This crucial challenge thereby motivates studies…
cs.LG2025
ResMoE: Space-efficient Compression of Mixture of Experts LLMs via Residual Restoration
Mengting Ai, Tianxin Wei, Yifan Chen +7
Mixture-of-Experts (MoE) Transformer, the backbone architecture of multiple phenomenal language models, leverages sparsity by activating only a fraction of model parameters for eac…