English
Related papers

Related papers: Any Data, Any Time, Anywhere: Global Data Access f…

200 papers

Science Data Systems (SDS) handle science data from acquisition through processing to distribution. They are deployed in the Cloud today, and the efficiency of Cloud instance utilization is critical to success. Conventional SDS are unable…

Distributed, Parallel, and Cluster Computing · Computer Science 2021-12-21 Lei Pan , Twinkle Jain

Today's astronomical projects need computational systems capable to store and analyze large amounts of scientific data, to effectively share data with other research Institutes and to easily implement information services to present data…

Astrophysics · Physics 2007-05-23 G. Calderone , L. Nicastro

Publicly available data from open sources (e.g., United States Census Bureau (Census), World Health Organization (WHO), Intergovernmental Panel on Climate Change (IPCC)) are vital resources for policy makers, students and researchers across…

The HEP community is approaching an era were the excellent performances of the particle accelerators in delivering collision at high rate will force the experiments to record a large amount of information. The growing size of the datasets…

Data is collected everywhere in our increasingly instrumented world and people are increasingly wanting to access this data from anywhere in it. This kind of anywhere & everywhere data present new challenges and opportunities for…

Human-Computer Interaction · Computer Science 2025-11-18 Niklas Elmqvist

The accumulation of a large amount of new experimental data at an impressive rate at present and future collider experiments has led to important questions concerning data storage and organization, their public access and usability, as well…

High Energy Physics - Phenomenology · Physics 2019-07-30 Andrea Ceccarelli , Andrea Cioni , Maria Vittoria Garzelli , Piergiulio Lenzi , Laura Redapi

The theoretical foundations of a new model and paradigm (called TIE) for data storage and access are introduced. Associations between data elements are stored in a single Matrix table, which is usually kept entirely in RAM for quick access.…

Databases · Computer Science 2007-05-23 Jerzy Lewak

Scientific experiments and modern applications are generating large amounts of data every day. Most organizations utilize In-house servers or Cloud resources to manage application data and workload. The traditional database management…

Databases · Computer Science 2025-06-17 Mayank Patel , Minal Bhise

Influenced by the advances in data and computing, the scientific practice increasingly involves machine learning and artificial intelligence driven methods which requires specialized capabilities at the system-, science- and service-level…

Distributed, Parallel, and Cluster Computing · Computer Science 2022-11-15 Ilkay Altintas , Ismael Perez , Dmitry Mishin , Adrien Trouillaud , Christopher Irving , John Graham , Mahidhar Tatineni , Thomas DeFanti , Shawn Strande , Larry Smarr , Michael L. Norman

Large High Energy Physics (HEP) experiments adopted a distributed computing model more than a decade ago. WLCG, the global computing infrastructure for LHC, in partnership with the US Open Science Grid, has achieved data management at the…

The CMS (Compact Muon Solenoid) experiment is one of the two large general-purpose particle physics detectors built at the LHC (Large Hadron Collider) at CERN in Geneva, Switzerland. The diverse collaboration combined with a highly…

Popular Physics · Physics 2011-10-04 Sudhir Malik , Kati Lassila-Perini

Data from high-energy physics (HEP) experiments are collected with significant financial and human effort and are in many cases unique. At the same time, HEP has no coherent strategy for data preservation and re-use, and many important and…

High Energy Physics - Experiment · Physics 2015-05-27 David M. South

For the past decade, HENP experiments have been heading towards a distributed computing model in an effort to concurrently process tasks over enormous data sets that have been increasing in size as a function of time. In order to optimize…

Distributed, Parallel, and Cluster Computing · Computer Science 2015-05-13 Michal Zerola , Jérôme Lauret , Roman Barták , Michal Šumbera

This is a personal perspective on data sharing in the context of public data releases suitable for generic analysis. These open data can be a powerful tool for expanding the science of high energy physics, but care must be taken in when,…

High Energy Physics - Phenomenology · Physics 2022-08-18 Benjamin Nachman

All Control Systems that grow to any size have a variety of data that are stored in different formats on different nodes in the network. Examples include sensor value and status, archived sensor data, device oriented support data and…

Accelerator Physics · Physics 2007-05-23 Matthias Clausen , Ron MacKenzie , Robert Sass , Kenneth Underwood , Greg White

The continuous growth of data production in almost all scientific areas raises new problems in data access and management, especially in a scenario where the end-users, as well as the resources that they can access, are worldwide…

Distributed, Parallel, and Cluster Computing · Computer Science 2022-08-16 Tommaso Tedeschi , Diego Ciangottini , Marco Baioletti , Valentina Poggioni , Daniele Spiga , Loriano Storchi , Mirco Tracolli

Scientific workflows have become highly heterogenous, leveraging distributed facilities such as High Performance Computing (HPC), Artificial Intelligence (AI), Machine Learning (ML), scientific instruments (data-driven pipelines) and edge…

Distributed, Parallel, and Cluster Computing · Computer Science 2024-11-27 Sadaf R. Alam , Christopher Woods , Matt Williams , Dave Moore , Isaac Prior , Ethan Williams , Anna Price , James Womack , Simon McIntosh-Smith , Fan Yang-Turner , Matt Pryor , Ilja Livenson

ESA Gaia mission is producing the more accurate source catalogue in astronomy up to now. That represents a challenge on the archiving area to make accessible this information to the astronomers in an efficient way. Also, new astronomical…

Exploratory data analysis tools must respond quickly to a user's questions, so that the answer to one question (e.g. a visualized histogram or fit) can influence the next. In some SQL-based query systems used in industry, even very large…

Distributed, Parallel, and Cluster Computing · Computer Science 2017-11-09 Jim Pivarski , David Lange , Thanat Jatuphattharachat