106 citations · 142 across the 8 of their papers we have counts for
7 papers · 1 filter
Farview: Disaggregated Memory with Operator Off-loading for Database Engines
Dario Korolija, Dimitrios Koutsoukos, Kimberly Keeton +3
Cloud deployments disaggregate storage from compute, providing more flexibility to both the storage and compute layers. In this paper, we explore disaggregation by taking it one st…
Evaluating Query Languages and Systems for High-Energy Physics Data [Extended Version]
Dan Graur, Ingo Müller, Mason Proffitt +3
In the domain of high-energy physics (HEP), query languages in general and SQL in particular have found limited acceptance. This is surprising since HEP data analysis matches the S…
The Collection Virtual Machine: An Abstraction for Multi-Frontend Multi-Backend Data Analysis
Ingo Müller, Renato Marroquín, Dimitrios Koutsoukos +3
Getting the best performance from the ever-increasing number of hardware platforms has been a recurring challenge for data processing systems. In recent years, the advent of data s…
Lambada: Interactive Data Analytics on Cold Data using Serverless Cloud Infrastructure
Ingo Müller, Renato Marroquín, Gustavo Alonso
The promise of ultimate elasticity and operational simplicity of serverless computing has recently lead to an explosion of research in this area. In the context of data analytics,…
Rumble: Data Independence for Large Messy Data Sets
Ingo Müller, Ghislain Fourny, Stefan Irimescu +2
This paper introduces Rumble, a query execution engine for large, heterogeneous, and nested collections of JSON objects built on top of Apache Spark. While data sets of this type a…
Pay One, Get Hundreds for Free: Reducing Cloud Costs through Shared Query Execution
Renato Marroquín, Ingo Müller, Darko Makreshanski +1
Cloud-based data analysis is nowadays common practice because of the lower system management overhead as well as the pay-as-you-go pricing model. The pricing model, however, is not…