The Rubin Observatory's Data Butler is designed to allow data file location and file formats to be abstracted away from the people writing the science pipeline algorithms. The Butler works in conjunction with the workflow graph builder to allow pipelines to be constructed from the algorithmic tasks. These pipelines can be executed at scale using object stores and multi-node clusters, or on a laptop using a local file system. The Butler and pipeline system are now in daily use during Rubin construction and early operations.
Cite
@article{arxiv.2206.14941,
title = {The Vera C. Rubin Observatory Data Butler and Pipeline Execution System},
author = {Tim Jenness and James F. Bosch and Nate B. Lust and Nathan M. Pease and Michelle Gower and Mikolaj Kowalik and Gregory P. Dubois-Felsmann and Fritz Mueller and Pim Schellart},
journal= {arXiv preprint arXiv:2206.14941},
year = {2022}
}
Comments
14 pages, 3 figures, submitted to Proc SPIE 12189, "Software and Cyberinfrastructure for Astronomy VII", Montreal, CA, July 2022