1 citations · 1 across the 1 of their papers we have counts for
2 papers
cs.DB2020★ 1 cited
Scalable Data Discovery Using Profiles
Javier Flores, Sergi Nadal, Oscar Romero
We study the problem of discovering joinable datasets at scale. This is, how to automatically discover pairs of attributes in a massive collection of independent, heterogeneous dat…
cs.DC2018
A Cost-based Storage Format Selector for Materialization in Big Data Frameworks
Rana Faisal Munir, Alberto Abelló, Oscar Romero +2
Modern big data frameworks (such as Hadoop and Spark) allow multiple users to do large-scale analysis simultaneously. Typically, users deploy Data-Intensive Workflows (DIWs) for th…