22 citations · 22 across the 3 of their papers we have counts for
3 papers
stat.ME2024★ 22 cited
A Selective Review on Statistical Methods for Massive Data Computation: Distributed Computing, Subsampling, and Minibatch Techniques
Xuetong Li, Yuan Gao, Hong Chang +11
This paper presents a selective review of statistical computation methods for massive data analysis. A huge amount of statistical methods for massive data computation have been rap…
stat.ME2023
CluBear: A Subsampling Package for Interactive Statistical Analysis with Massive Data on A Single Machine
Ke Xu, Yingqiu Zhu, Yijing Liu +1
This article introduces CluBear, a Python-based open-source package for interactive massive data analysis. The key feature of CluBear is that it enables users to conduct convenient…
cs.LG2023
Improved Naive Bayes with Mislabeled Data
Qianhan Zeng, Yingqiu Zhu, Xuening Zhu +5
Labeling mistakes are frequently encountered in real-world applications. If not treated well, the labeling mistakes can deteriorate the classification performances of a model serio…