47 citations · 47 across the 2 of their papers we have counts for
2 papers
cs.CL2024
AquilaMoE: Efficient Training for MoE Models with Scale-Up and Scale-Out Strategies
Bo-Wen Zhang, Liangdong Wang, Ye Yuan +24
In recent years, with the rapid application of large language models across various fields, the scale of these models has gradually increased, and the resources required for their…
cs.CV2021★ 47 cited
Self-Supervised Pre-Training for Transformer-Based Person Re-Identification
Hao Luo, Pichao Wang, Yi Xu +5
Transformer-based supervised pre-training achieves great performance in person re-identification (ReID). However, due to the domain gap between ImageNet and ReID datasets, it usual…