Related papers: The Vera C. Rubin Observatory Data Butler and Pipe…
Informative planning seeks a sequence of actions that guide the robot to collect the most informative data to build a large-scale environmental model or learn a dynamical system. Existing work in informative planning mainly focuses on…
Achieving a percentage-level precision measurement of the Coherent Elastic Neutrino Nucleus Scattering (CE{\nu}NS) spectrum requires a robust data processing pipeline which can be characterised with great precision. To fulfil this goal we…
The increasing adoption of low-cost environmental sensors and AI-enabled applications has accelerated the demand for scalable and resilient data infrastructures, particularly in data-scarce and resource-constrained regions. This paper…
We present the first public release of ShapePipe, an open-source and modular weak-lensing measurement, analysis, and validation pipeline written in Python. We describe the design of the software and justify the choices made. We provide a…
Cloud infrastructure supports the efficient operation of data pipelines regarding requirements like cost, speed, and resource utilization. We present an integrated view of optimization opportunities for cloud-based data pipelines by…
SOFIA presents a number of interesting challenges for the development of a data reduction environment which, at its initial phase, will have to incorporate pipelines from seven different instruments. Therefore, the SOFIA data reduction…
Machine learning pipeline potentially consists of several stages of operations like data preprocessing, feature engineering and machine learning model training. Each operation has a set of hyper-parameters, which can become irrelevant for…
In the multi-messenger era, astronomical projects share information about transients phenomena issuing science alerts to the Scientific Community through different communications networks. This coordination is mandatory to understand the…
Robotic wide-field time-domain surveys, such as the Zwicky Transient Facility and the Asteroid Terrestrial-impact Last Alert System, capture dozens of transients each night. The workflows for discovering and classifying transients in survey…
As the volume of data available from sensor-enabled devices such as vehicles expands, it is increasingly hard for companies to make informed decisions about the cost of capturing, processing, and storing the data from every device. Business…
The Kepler Mission was designed to identify and characterize transiting planets in the Kepler Field of View and to determine their occurrence rates. Emphasis was placed on identification of Earth-size planets orbiting in the Habitable Zone…
The aim of this paper is to develop an approach to visualizations that benefits from distributed computing. Three schemes of process distribution are considered: parallel, pipeline, and expanding pipeline computations. Expanding pipeline…
The Pulsar Virtual Observatory will provide a means for scientists in all fields to access and analyze the large data sets stored in pulsar surveys without specific knowledge about the data or the processing mechanisms. This is achieved by…
Performance of the Level-2 pipeline, which translates the UVIT data created by the ISRO's ground segment processing systems (Level-1) into astronomer ready scientific data products, is described. This pipeline has evolved significantly from…
The digital transformation of production requires new methods of data integration and storage, as well as decision making and support systems that work vertically and horizontally throughout the development, production, and use cycle. In…
Kuiper belt objects smaller than a few kilometers are difficult to observe directly. They can be detected when they randomly occult a background star. Close to the ecliptic plane, each star is occulted once every tens of thousands of hours,…
The evaluation of informative path planning algorithms for autonomous vehicles is often hindered by fragmented execution pipelines and limited transferability between simulation and real-world deployment. This paper introduces a unified…
We present three Virtual Observatory tools developed at the ATNF for the storage, processing and visualisation of ATCA data. These are the Australia Telescope Online Archive, a prototype data reduction pipeline, and the Remote Visualisation…
Kepler's primary mission is a search for earth-size exoplanets in the habitable zone of late-type stars using the transit method. To effectively accomplish this mission, Kepler orbits the Sun and stares nearly continuously at one…
Data scientists develop ML pipelines in an iterative manner: they repeatedly screen a pipeline for potential issues, debug it, and then revise and improve its code according to their findings. However, this manual process is tedious and…