2 papers
cs.SE2024
PerfGen: Automated Performance Benchmark Generation for Big Data Analytics
Jiyuan Wang, Jason Teoh, Muhammand Ali Gulza +2
Many symptoms of poor performance in big data analytics such as computational skews, data skews, and memory skews are input dependent. However, due to the lack of inputs that can t…
cs.SE2021
Efficient Fuzz Testing for Apache Spark Using Framework Abstraction
Qian Zhang, Jiyuan Wang, Muhammad Ali Gulzar +2
The emerging data-intensive applications are increasingly dependent on data-intensive scalable computing (DISC) systems, such as Apache Spark, to process large data. Despite their…