24 citations · 24 across the 2 of their papers we have counts for
2 papers
cs.DC2023★ 24 cited
Towards General and Efficient Online Tuning for Spark
Yang Li, Huaijun Jiang, Yu Shen +8
The distributed data analytic system -- Spark is a common choice for processing massive volumes of heterogeneous data, while it is challenging to tune its parameters to achieve hig…
cs.LG2023
Rover: An online Spark SQL tuning service via generalized transfer learning
Yu Shen, Xinyuyang Ren, Yupeng Lu +6
Distributed data analytic engines like Spark are common choices to process massive data in industry. However, the performance of Spark SQL highly depends on the choice of configura…