1 citations · 1 across the 1 of their papers we have counts for
2 papers
cs.DB2024
FREYJA: Efficient Join Discovery in Data Lakes
Marc Maynou, Sergi Nadal, Raquel Panadero +3
Data lakes are massive repositories of raw and heterogeneous data, designed to meet the requirements of modern data storage. Nonetheless, this same philosophy increases the complex…
cs.DB2020★ 1 cited
Scalable Data Discovery Using Profiles
Javier Flores, Sergi Nadal, Oscar Romero
We study the problem of discovering joinable datasets at scale. This is, how to automatically discover pairs of attributes in a massive collection of independent, heterogeneous dat…