3 citations · 3 across the 3 of their papers we have counts for
3 papers
cs.DB2026
Index-Assisted Stratified Sampling for Online Aggregation
Yunnan Yu, Zhuoyue Zhao
Ad-hoc queries over frequently updated data in a flat schema are common in real-time data analysis applications and often require very low latency. Online aggregation can achieve s…
cs.DB2025
SABER: A SQL-Compatible Semantic Document Processing System Based on Extended Relational Algebra
Changjae Lee, Zhuoyue Zhao, Jinjun Xiong
The emergence of large-language models (LLMs) has enabled a new class of semantic data processing systems (SDPSs) to support declarative queries against unstructured documents. Exi…
cs.DB2017★ 3 cited
InferSpark: Statistical Inference at Scale
Zhuoyue Zhao, Jialing Pei, Eric Lo +2
The Apache Spark stack has enabled fast large-scale data processing. Despite a rich library of statistical models and inference algorithms, it does not give domain users the abilit…