4 papers · 1 filter
Auto-Relate: A Unified Approach to Discovering Reliable Functional Relationships Leveraging Statistical Tests
Ziyan Han, Yeye He, Shuyuan Kang +8
Tables in spreadsheets, computational notebooks, and databases often contain rich inter-column relationships. Yet these relationships are typically implicit and are often lost when…
Efficient Query Rewrite Rule Discovery via Standardized Enumeration and Learning-to-Rank(extend)
Yuan Zhang, Yuxing Chen, Yuekun Yu +5
Query rewriting is essential for database performance optimization, but existing automated rule enumeration methods suffer from exponential search spaces, severe redundancy, and po…
Privacy-Enhanced Database Synthesis for Benchmark Publishing (Technical Report)
Yunqing Ge, Jianbin Qin, Shuyuan Zheng +7
Benchmarking is crucial for evaluating a DBMS, yet existing benchmarks often fail to reflect the varied nature of user workloads. As a result, there is increasing momentum toward c…
Automatic String Data Validation with Pattern Discovery
Xinwei Lin, Jing Zhao, Peng Di +7
In enterprise data pipelines, data insertions occur periodically and may impact downstream services if data quality issues are not addressed. Typically, such problems can be invest…