4 papers · 1 filter
Auto-Relate: A Unified Approach to Discovering Reliable Functional Relationships Leveraging Statistical Tests
Ziyan Han, Yeye He, Shuyuan Kang +8
Tables in spreadsheets, computational notebooks, and databases often contain rich inter-column relationships. Yet these relationships are typically implicit and are often lost when…
Efficient Query Rewrite Rule Discovery via Standardized Enumeration and Learning-to-Rank(extend)
Yuan Zhang, Yuxing Chen, Yuekun Yu +5
Query rewriting is essential for database performance optimization, but existing automated rule enumeration methods suffer from exponential search spaces, severe redundancy, and po…
Automatic String Data Validation with Pattern Discovery
Xinwei Lin, Jing Zhao, Peng Di +7
In enterprise data pipelines, data insertions occur periodically and may impact downstream services if data quality issues are not addressed. Typically, such problems can be invest…
Privacy-Enhanced Database Synthesis for Benchmark Publishing (Technical Report)
Yunqing Ge, Jianbin Qin, Shuyuan Zheng +7
Benchmarking is crucial for evaluating a DBMS, yet existing benchmarks often fail to reflect the varied nature of user workloads. As a result, there is increasing momentum toward c…