1 citations · 2 across the 4 of their papers we have counts for
4 papers
Arena Learning: Build Data Flywheel for LLMs Post-training via Simulated Chatbot Arena
Haipeng Luo, Qingfeng Sun, Can Xu +6
Assessing the effectiveness of large language models (LLMs) presents substantial challenges. The method of conducting human-annotated battles in an online Chatbot Arena is a highly…
Contrastive Learning with Negative Sampling Correction
Lu Wang, Chao Du, Pu Zhao +8
As one of the most effective self-supervised representation learning methods, contrastive learning (CL) relies on multiple negative pairs to contrast against each positive pair. In…
Diffusion-based Time Series Data Imputation for Microsoft 365
Fangkai Yang, Wenjie Yin, Lu Wang +10
Reliability is extremely important for large-scale cloud systems like Microsoft 365. Cloud failures such as disk failure, node failure, etc. threaten service reliability, resulting…
Robust Positive-Unlabeled Learning via Noise Negative Sample Self-correction
Zhangchi Zhu, Lu Wang, Pu Zhao +7
Learning from positive and unlabeled data is known as positive-unlabeled (PU) learning in literature and has attracted much attention in recent years. One common approach in PU lea…