9 citations · 16 across the 5 of their papers we have counts for
5 papers · 1 filter
TSEXPLAIN: Explaining Aggregated Time Series by Surfacing Evolving Contributors
Yiru Chen, Silu Huang
Aggregated time series are generated effortlessly everywhere, e.g., "total confirmed covid-19 cases since 2019" and "total liquor sales over time." Understanding "how" and "why" th…
OrpheusDB: Bolt-on Versioning for Relational Databases
Silu Huang, Liqi Xu, Jialin Liu +2
Data science teams often collaboratively analyze datasets, generating dataset versions at each stage of iterative exploration and analysis. There is a pressing need for a system th…
Finding Multiple New Optimal Locations in a Road Network
Ruifeng Liu, Ada WaiChee Fu, Zitong Chen +2
We study the problem of optimal location querying for location based services in road networks, which aims to find locations for new servers or facilities. The existing optimal sol…
Towards a unified query language for provenance and versioning
Amit Chavan, Silu Huang, Amol Deshpande +3
Organizations and teams collect and acquire data from various sources, such as social interactions, financial transactions, sensor data, and genome sequencers. Different teams in a…
Principles of Dataset Versioning: Exploring the Recreation/Storage Tradeoff
Souvik Bhattacherjee, Amit Chavan, Silu Huang +2
The relative ease of collaborative data science and analysis has led to a proliferation of many thousands or millions of of the same datasets in many scientific and comm…