English
Related papers

Related papers: A Bayesian quantification of consistency in correl…

200 papers

We study the problem of linear feature selection when features are highly correlated. Such settings pose two fundamental challenges. First, how should model similarity be defined? Simply counting features in common can be misleading: two…

Methodology · Statistics 2026-03-24 Xiaozhu Zhang , Jacob Bien , Armeen Taeb

While observational data are routinely used to estimate causal effects of biomedical treatments, doing so requires special methods to adjust for observed confounding. These methods invariably rely on untestable statistical and causal…

Methodology · Statistics 2026-03-02 Arman Oganisian

We derive an exact and efficient Bayesian regression algorithm for piecewise constant functions of unknown segment number, boundary location, and levels. It works for any noise and segment level prior, e.g. Cauchy which can handle outliers.…

Statistics Theory · Mathematics 2007-06-13 Marcus Hutter

Correlated component analysis as proposed by Dmochowski et al. (2012) is a tool for investigating brain process similarity in the responses to multiple views of a given stimulus. Correlated components are identified under the assumption…

Machine Learning · Statistics 2018-02-08 Simon Kamronn , Andreas Trier Poulsen , Lars Kai Hansen

Estimating time-varying correlation matrices is challenging because existing methods may adapt slowly to structural changes, impose insufficient regularization, or produce diffuse posterior uncertainty. In moderate dimensions, an additional…

Methodology · Statistics 2026-05-11 Daniel Andrew Coulson , David S. Matteson , Martin T. Wells

Statistical modeling and inference problems with sample sizes substantially smaller than the number of available covariates are challenging. Chakraborty et al. (2012) did a full hierarchical Bayesian analysis of nonlinear regression in such…

Machine Learning · Statistics 2022-02-14 Xiao Fang , Malay Ghosh

We use numerical simulations of ray tracing through N-body simulations to investigate weak lensing by large-scale structure. These are needed for testing the analytic predictions of two-point correlators, to set error estimates on them and…

Astrophysics · Physics 2007-05-23 Bhuvnesh Jain , Uros Seljak , Simon White

We empirically show that Bayesian inference can be inconsistent under misspecification in simple linear regression problems, both in a model averaging/selection and in a Bayesian ridge regression setting. We use the standard linear model,…

Statistics Theory · Mathematics 2018-10-30 Peter Grünwald , Thijs van Ommen

We analyse three public cosmic shear surveys; the Kilo-Degree Survey (KiDS-450), the Dark Energy Survey (DES-SV) and the Canada France Hawaii Telescope Lensing Survey (CFHTLenS). Adopting the COSEBIs statistic to cleanly and completely…

We present a comparative study of the accuracy and precision of correlation function methods and full-field inference in cosmological data analysis. To do so, we examine a Bayesian hierarchical model that predicts log-normal fields and…

Cosmology and Nongalactic Astrophysics · Physics 2021-09-07 Florent Leclercq , Alan Heavens

We present the methodology for a joint cosmological analysis of weak gravitational lensing from the fourth data release of the ESO Kilo-Degree Survey (KiDS-1000) and galaxy clustering from the partially overlapping BOSS and 2dFLenS surveys.…

We use the halo model of clustering to compute two- and three-point correlation functions for weak lensing, and apply them in a new statistical technique to measure properties of massive halos. We present analytical results on the eight…

Astrophysics · Physics 2009-11-07 Masahiro Takada , Bhuvnesh Jain

Bayesian inference is now a leading technique for reconstructing phylogenetic trees from aligned sequence data. In this short note, we formally show that the maximum posterior tree topology provides a statistically consistent estimate of a…

Populations and Evolution · Quantitative Biology 2013-07-12 Mike Steel

We perform a combined analysis of cosmic shear tomography, galaxy-galaxy lensing tomography, and redshift-space multipole power spectra (monopole and quadrupole) using 450 deg$^2$ of imaging data by the Kilo Degree Survey (KiDS) overlapping…

Measuring dataset similarity is fundamental in machine learning, particularly for transfer learning and domain adaptation. In the context of supervised learning, most existing approaches quantify similarity of two data sets based on their…

Machine Learning · Statistics 2026-04-22 Shudong Sun , Hao Helen Zhang , Joseph C Watkins

Building a machine learning solution in real-life applications often involves the decomposition of the problem into multiple models of various complexity. This has advantages in terms of overall performance, better interpretability of the…

Artificial Intelligence · Computer Science 2020-05-27 Bashar Awwad Shiekh Hasan , Kate Kelly

Accurate comparisons between theoretical models and experimental data are critical for scientific progress. However, inferred physical model parameters can vary significantly with the chosen physics model, highlighting the importance of…

High Energy Physics - Phenomenology · Physics 2025-10-27 Sunil Jaiswal , Chun Shen , Richard J. Furnstahl , Ulrich Heinz , Matthew T. Pratola

Learning a good distance metric in feature space potentially improves the performance of the KNN classifier and is useful in many real-world applications. Many metric learning algorithms are however based on the point estimation of a…

Computer Vision and Pattern Recognition · Computer Science 2016-04-11 Dong Wang , Xiaoyang Tan

In modern data analysis, nonparametric measures of discrepancies between random variables are particularly important. The subject is well-studied in the frequentist literature, while the development in the Bayesian setting is limited where…

Methodology · Statistics 2022-01-25 Qinyi Zhang , Veit Wild , Sarah Filippi , Seth Flaxman , Dino Sejdinovic