586 citations · 654 across the 3 of their papers we have counts for
Showing 2018Show all
2 papers · 1 filter
astro-ph.IM2018
Analyzing billion-objects catalog interactively: Apache Spark for physicists
S. Plaszczynski, J. Peloton, C. Arnault +1
Apache Spark is a Big Data framework for working on large distributed datasets. Although widely used in the industry, it remains rather limited in the academic community or often r…
astro-ph.IM2018
FITS Data Source for Apache Spark
Julien Peloton, Christian Arnault, Stéphane Plaszczynski
We investigate the performance of Apache Spark, a cluster computing framework, for analyzing data from future LSST-like galaxy surveys. Apache Spark attempts to address big data pr…