Related papers: Computational AstroStatistics: Fast and Efficient …
Astronomical observations typically provide three-dimensional maps, encoding the distribution of the observed flux in (1) the two angles of the celestial sphere and (2) energy/frequency. An important task regarding such maps is to…
The development of fast and accurate methods of photometric redshift estimation is a vital step towards being able to fully utilize the data of next-generation surveys within precision cosmology. In this paper we apply a specific approach…
New developments in data processing and visualization are being made in preparation for upcoming radioastronomical surveys planned with the Square Kilometre Array (SKA) and its precursors. A major goal is enabling extraction of science…
Innovative developments in data processing, archiving, analysis, and visualization are nowadays unavoidable to deal with the data deluge expected in next-generation facilities for radio astronomy, such as the Square Kilometre Array (SKA)…
This work investigates symbolic regression (SR) as an interpretable alternative to black-box machine learning for the classification of stars, galaxies, and quasars in the Sloan Digital Sky Survey Data Release 17 (SDSS DR17). We conduct a…
We present a new algorithm called 'Fast Integrated Spectra Analyzer" (FISA) that permits fast and reasonably accurate age and reddening determinations for small angular diameter open clusters by using their integrated spectra in the…
We present a data-driven technique to analyze multifrequency images from upcoming cosmological surveys mapping large sky area. Using full information from the data at the two-point level, our method can simultaneously constrain the…
We present an algorithm for the fast computation of the general $N$-point spatial correlation functions of any discrete point set embedded within an Euclidean space of $\mathbb{R}^n$. Utilizing the concepts of kd-trees and graph databases,…
The analysis and an efficient scientific exploration of the Digital Palomar Observatory Sky Survey (DPOSS) represents a major technical challenge. The input data set consists of 3 Terabytes of pixel information, and contains a few billion…
In Astrophysics, the identification of candidate Globular Clusters through deep, wide-field, single band HST images, is a typical data analytics problem, where methods based on Machine Learning have revealed a high efficiency and…
We present an algorithm capable of detecting diffuse, dim sources of any size in an astronomical image. These sources often defeat traditional methods for source finding, which expand regions around points of high intensity. Extended…
We present recent results from the Laboratory for Cosmological Data Mining (http://lcdm.astro.uiuc.edu) at the National Center for Supercomputing Applications (NCSA) to provide robust classifications and photometric redshifts for objects in…
The Sloan Digital Sky Survey (SDSS) is collecting photometry and intermediate resolution spectra for ~ 10**5 stars in the thick-disk and stellar halo of the Milky Way. This massive dataset can be used to infer the properties of the stars…
The luminosity changes of most types of variable stars are correlated in the different wavelengths, and these correlations may be exploited for several purposes: for variability detection, for distinction of microvariability from noise, for…
This review outlines concepts of mathematical statistics, elements of probability theory, hypothesis tests and point estimation for use in the analysis of modern astronomical data. Least squares, maximum likelihood, and Bayesian approaches…
The Sloan Digital Sky Survey (SDSS) currently provides by far the largest homogeneous sample of intermediate signal-to-noise ratio (S/N) optical spectra of galaxies and quasars. The fully automated SDSS spectroscopic reduction pipeline has…
The development of search algorithms for gravitational wave sources in the LISA data stream is currently a very active area of research. It has become clear that not only does difficulty lie in searching for the individual sources, but in…
We have undertaken a dedicated program of automatic source classification in the WISE database merged with SuperCOSMOS scans, comprehensively identifying galaxies, quasars and stars on most of the unconfused sky. We use the Support Vector…
Generation of science-ready data from processed data products is one of the major challenges in next-generation radio continuum surveys with the Square Kilometre Array (SKA) and its precursors, due to the expected data volume and the need…
In area and depth, the Pan-STARRS1 (PS1) 3$\pi$ survey is unique among many-epoch, multi-band surveys and has enormous potential for all-sky identification of variable sources. PS1 has observed the sky typically seven times in each of its…