activity
20122026
most citedCrowdER: Crowdsourcing Entity Resolution

66 citations · 138 across the 9 of their papers we have counts for

collaborators
Showing cs.DBShow all

11 papers · 1 filter

cs.DB2026

The Case for Text-to-SQL Friendly Logical Database Design

Shi Heng Zhang, Zhengjie Miao, Jiannan Wang

Logical database design has traditionally optimized database schemas, including tables, columns, keys, constraints, and views, for correctness, integrity, and human-written applica…

cs.DB2026

PrepBench: How Far Are We from Natural-Language-Driven Data Preparation?

Jingzhe Xu, Rui Wang, Jiannan Wang +1

Data preparation is a central and time-consuming stage in data analysis workflows. Traditionally, commercial tools have relied on graphical user interfaces (GUIs) to simplify data…

cs.DB202129 cited

DataPrep.EDA: Task-Centric Exploratory Data Analysis for Statistical Modeling in Python

Jinglin Peng, Weiyuan Wu, Brandon Lockhart +6

Exploratory Data Analysis (EDA) is a crucial step in any data science project. However, existing Python libraries fall short in supporting data scientists to complete common EDA ta…

cs.DB2021

Explaining Inference Queries with Bayesian Optimization

Brandon Lockhart, Jinglin Peng, Weiyuan Wu +2

Obtaining an explanation for an SQL query result can enrich the analysis experience, reveal data errors, and provide deeper insight into the data. Inference query explanation seeks…

cs.DB2020

Are We Ready For Learned Cardinality Estimation?

Xiaoying Wang, Changbo Qu, Weiyuan Wu +2

Cardinality estimation is a fundamental but long unresolved problem in query optimization. Recently, multiple papers from different research groups consistently report that learned…

cs.DB202036 cited

Complaint-driven Training Data Debugging for Query 2.0

Weiyuan Wu, Lampros Flokas, Eugene Wu +1

As the need for machine learning (ML) increases rapidly across all industry sectors, there is a significant interest among commercial database providers to support "Query 2.0", whi…