Related papers: Reliability Correction is Key for Robust Kepler Oc…
In many applications, accurate class probability estimates are required, but many types of models produce poor quality probability estimates despite achieving acceptable classification accuracy. Even though probability calibration has been…
Machine learning-supported decisions, such as ordering diagnostic tests or determining preventive custody, often require converting probabilistic forecasts into binary classifications. We adopt a consequentialist perspective from decision…
NASA's Kepler Space Telescope was designed to determine the frequency of Earth-sized planets orbiting Sun-like stars, but these planets are on the very edge of the mission's detection sensitivity. Accurately determining the occurrence rate…
The problem of model selection is inevitable in an increasingly large number of applications involving partial theoretical knowledge and vast amounts of information, like in medicine, biology or economics. The associated techniques are…
We present occurrence rates for rocky planets in the habitable zones (HZ) of main-sequence dwarf stars based on the Kepler DR25 planet candidate catalog and Gaia-based stellar properties. We provide the first analysis in terms of…
The Kepler Mission has found thousands of planetary candidates with radii between 1 and 4 R$_\oplus$. These planets have no analogues in our own Solar System, providing an unprecedented opportunity to understand the range and distribution…
Doppler planet searches have discovered that giant planets follow orbits with a wide range of orbital eccentricities, revolutionizing theories of planet formation. The discovery of hundreds of exoplanet candidates by NASA's Kepler mission…
We constrain the densities of Earth- to Neptune-size planets around very cool (Te =3660-4660K) Kepler stars by comparing 1202 Keck/HIRES radial velocity measurements of 150 nearby stars to a model based on Kepler candidate planet radii and…
The Kepler Mission was launched on March 6, 2009 to perform a photometric survey of more than 100,000 dwarf stars to search for Earth-size planets with the transit technique. The reliability of the resulting planetary candidate list relies…
Sigma clipping is commonly used in astronomy for outlier rejection, but the number of standard deviations beyond which one should clip data from a sample ultimately depends on the size of the sample. Chauvenet rejection is one of the…
Galaxy distances and derived radial peculiar velocity catalogs constitute valuable datasets to study the dynamics of the Local Universe. However, such catalogs suffer from biases whose effects increase with the distance. Malmquist biases…
One bottleneck for the exploitation of data from the $Kepler$ mission for stellar astrophysics and exoplanet research has been the lack of precise radii and evolutionary states for most of the observed stars. We report revised radii of…
We carry out an independent search of Kepler photometry for small transiting planets with sizes 0.5--8.0 times that of Earth and orbital periods between 5 and 50 days, with the goal of measuring the fraction of stars harboring such planets.…
Kepler will monitor enough stars that it is likely to detect single transits of planets with periods longer than the mission lifetime. We show that by combining the Kepler photometry of such transits with precise radial velocity (RV)…
Model selection is a cornerstone of statistical inference, where information criteria are widely employed to balance model fit and complexity. However, classical likelihood-based criteria are often highly sensitive to contamination,…
This book chapter introduces regression approaches and regression adjustment for Approximate Bayesian Computation (ABC). Regression adjustment adjusts parameter values after rejection sampling in order to account for the imperfect match…
We perform a search for transiting planets in the NASA K2 observations of the globular cluster (GC) M4. This search is sensitive to larger orbital periods ($P\lesssim 35$ days, compared to the previous best of $P\lesssim 16$ days) and, at…
We propose a generic numerical measure of the inconsistency of a database with respect to a set of integrity constraints. It is based on an abstract repair semantics. In particular, an inconsistency measure associated to cardinality-repairs…
As in other estimation scenarios, likelihood based estimation in the normal mixture set-up is highly non-robust against model misspecification and presence of outliers (apart from being an ill-posed optimization problem). A robust…
Inferring the causal effect of a treatment on an outcome in an observational study requires adjusting for observed baseline confounders to avoid bias. However, adjusting for all observed baseline covariates, when only a subset are…