1 citations · 1 across the 3 of their papers we have counts for
5 papers
Exploring Erasure Coding Techniques for High Availability of Intermediate Data
Zhe Zhang, Brian Bockelman, Derek Weitzel +1
Scientific computing workflows generate enormous distributed data that is short-lived, yet critical for job completion time. This class of data is called intermediate data. A commo…
Trua: Efficient Task Replication for Flexible User-defined Availability in Scientific Grids
Zhe Zhang, Brian Bockelman, Derek Weitzel +3
Failure is inevitable in scientific computing. As scientific applications and facilities increase their scales over the last decades, finding the root cause of a failure can be ver…
ROOT I/O compression improvements for HEP analysis
Oksana Shadura, Brian Paul Bockelman, Philippe Canal +2
We overview recent changes in the ROOT I/O system, increasing performance and enhancing it and improving its interaction with other data analysis ecosystems. Both the newly introdu…
Speeding HEP Analysis with ROOT Bulk I/O
Brian Bockelman, Zhe Zhang, Oksana Shadura
Distinct HEP workflows have distinct I/O needs; while ROOT I/O excels at serializing complex C++ objects common to reconstruction, analysis workflows typically have simpler objects…
Fast Access to Columnar, Hierarchically Nested Data via Code Transformation
Jim Pivarski, Peter Elmer, Brian Bockelman +1
Big Data query systems represent data in a columnar format for fast, selective access, and in some cases (e.g. Apache Drill), perform calculations directly on the columnar data wit…