Related papers: Scaling pair count to next galaxy surveys
The analysis of Redshift-Space Distortions (RSD) within galaxy surveys provides constraints on the amplitude of peculiar velocities induced by structure growth, thereby allowing tests of General Relativity on extremely large scales. The…
We present a scalable, cloud-based science platform solution designed to enable next-to-the-data analyses of terabyte-scale astronomical tabular datasets. The presented platform is built on Amazon Web Services (over Kubernetes and S3…
This paper proposes a general adaptive procedure for budget-limited predictor design in high dimensions called two-stage Sampling, Prediction and Adaptive Regression via Correlation Screening (SPARCS). SPARCS can be applied to high…
The currently operating space missions, as well as those that will be launched in the near future, (will) deliver high-quality data for millions of stellar objects. Since the majority of stellar astrophysical applications still (at least…
Recently there has been an increase in the studies on time-series data mining specifically time-series clustering due to the vast existence of time-series in various domains. The large volume of data in the form of time-series makes it…
Close pairs of galaxies have been broadly studied in the literature as a way to understand galaxy interactions and mergers. In observations they are usually defined by setting a maximum separation in the sky and in velocity along the line…
Super-sample covariance (SSC) is the dominant source of statistical error on large scale structure (LSS) observables for both current and future galaxy surveys. In this work, we concentrate on the SSC of cluster counts, also known as sample…
We present the description of the project \texttt{SCORPIO}, a Python package for retrieving images and associated data of galaxy pairs based on their position, facilitating visual analysis and data collation of multiple archetypal systems.…
Weak gravitational lensing is a powerful probe for constraining cosmological parameters, but its success relies on accurate shear measurements. In this paper, we use image simulations to investigate how a joint analysis of high-resolution…
Recent works on crowd counting mainly leverage CNNs to count by regressing density maps, and have achieved great progress. In the density map, each person is represented by a Gaussian blob, and the final count is obtained from the…
Clusters of galaxies are the most massive objects in the Universe and mapping their location is an important astronomical problem. This paper describes an algorithm (based on statistical signal processing methods), a software architecture…
A robust measurement of the clustering amplitude of the sub-mm population of starburst galaxies requires large-area surveys (>> 1 deg^2). The largest-format arrays subtend only 10 arcmin^2 on the sky and hence scan-mapping is a necessary…
We introduce a highly efficient method for panoptic segmentation of large 3D point clouds by redefining this task as a scalable graph clustering problem. This approach can be trained using only local auxiliary tasks, thereby eliminating the…
Generating wide-area digital surface models (DSMs) requires registering a large number of individual, and partially overlapped DSMs. This presents a challenging problem for a typical registration algorithm, since when a large number of…
The next generation of proposed galaxy surveys will increase the number of galaxies with photometric redshifts by two orders of magnitude, drastically expanding both redshift range and detection threshold from the current state of the art.…
We present a numerically cheap approximation to super-sample covariance (SSC) of large scale structure cosmological probes, first in the case of angular power spectra. It necessitates no new elements besides those used for the prediction of…
We present the first results from GALAXY CRUISE, a community (or citizen) science project based on data from the Hyper Suprime-Cam Subaru Strategic Program (HSC-SSP). The current paradigm of galaxy evolution suggests that galaxies grow…
The immense amount of daily generated and communicated data presents unique challenges in their processing. Clustering, the grouping of data without the presence of ground-truth labels, is an important tool for drawing inferences from data.…
The ESA Euclid mission will provide high-quality imaging for about 1.5 billion galaxies. A software pipeline to automatically process and analyse such a huge amount of data in real time is being developed by the Science Ground Segment of…
Two new algorithms are described for matching two dimensional coordinate lists of point sources that are signifcantly faster than previous methods. By matching rarely occurring triangles (or more complex shapes) in the two lists, and by…