English
Related papers

Related papers: Occam's Ghost

200 papers

We introduce the concept of compressed convolution, a technique to convolve a given data set with a large number of non-orthogonal kernels. In typical applications our technique drastically reduces the effective number of computations. The…

Instrumentation and Methods for Astrophysics · Physics 2014-01-08 F. Elsner , B. D. Wandelt

Probabilistic Regression refers to predicting a full probability density function for the target conditional on the features. We present a nonparametric approach to this problem which combines base classifiers (typically gradient boosted…

Machine Learning · Computer Science 2022-10-31 Brian Lucena

The Waxman random graph is a generalisation of the simple Erd\H{o}s-R\'enyi or Gilbert random graph. It is useful for modelling physical networks where the increased cost of longer links means they are less likely to be built, and thus less…

Statistics Theory · Mathematics 2015-06-30 Matthew Roughan , Jonathan Tuke , Eric Parsonage

As in other estimation scenarios, likelihood based estimation in the normal mixture set-up is highly non-robust against model misspecification and presence of outliers (apart from being an ill-posed optimization problem). A robust…

Methodology · Statistics 2023-12-20 Soumya Chakraborty , Ayanendranath Basu , Abhik Ghosh

Detecting symmetry from data is a fundamental problem in signal analysis, providing insight into underlying structure and constraints. When data emerge as trajectories of dynamical systems, symmetries encode structural properties of the…

Machine Learning · Statistics 2025-10-21 Ziad Ghanem , Chang Hyunwoong , Preskella Mrad

We study the relationship between gradient-based optimization of parametric models (e.g., neural networks) and optimization of linear combinations of random features. Our main result shows that if a parametric model can be learned using…

Machine Learning · Computer Science 2025-05-16 Ari Karchmer , Eran Malach

In a circular deconvolution model we consider the fully data driven density estimation of a circular random variable where the density of the additive independent measurement error is unknown. We have at hand two independent iid samples,…

Statistics Theory · Mathematics 2021-02-02 Jan Johannes , Xavier Loizeau

Histograms are convenient non-parametric density estimators, which continue to be used ubiquitously. Summary quantities estimated from histogram-based probability density models depend on the choice of the number of bins. We introduce a…

Data Analysis, Statistics and Probability · Physics 2013-09-17 Kevin H. Knuth

Roughly, a metric space has padding parameter $\beta$ if for every $\Delta>0$, there is a stochastic decomposition of the metric points into clusters of diameter at most $\Delta$ such that every ball of radius $\gamma\Delta$ is contained in…

Data Structures and Algorithms · Computer Science 2025-04-02 Jonathan Conroy , Arnold Filtser

We consider a Bayesian approach to model selection in Gaussian linear regression, where the number of predictors might be much larger than the number of observations. From a frequentist view, the proposed procedure results in the penalized…

Statistics Theory · Mathematics 2010-09-14 Felix Abramovich , Vadim Grinshtein

Subsets of F_2^n that are eps-biased, meaning that the parity of any set of bits is even or odd with probability eps close to 1/2, are powerful tools for derandomization. A simple randomized construction shows that such sets exist of size…

Computational Complexity · Computer Science 2013-04-19 Cristopher Moore , Alexander Russell

The problem we concentrate on is as follows: given (1) a convex compact set $X$ in ${\mathbb{R}}^n$, an affine mapping $x\mapsto A(x)$, a parametric family $\{p_{\mu}(\cdot)\}$ of probability densities and (2) $N$ i.i.d. observations of the…

Statistics Theory · Mathematics 2009-08-24 Anatoli B. Juditsky , Arkadi S. Nemirovski

High throughput genetic sequencing arrays with thousands of measurements per sample and a great amount of related censored clinical data have increased demanding need for better measurement specific model selection. In this paper we…

Statistics Theory · Mathematics 2019-07-31 Jelena Bradic , Jianqing Fan , Jiancheng Jiang

In this paper, we will discuss how to generalize nonparametric density estimators to MLE parametric estimators. Basing on the Parzen window theory and using the advantages of probability amplitude of quantum theory, we model a nonlinear…

Statistics Theory · Mathematics 2008-11-13 Yeong-Shyeong Tsai

Modern neural network architectures typically have many millions of parameters and can be pruned significantly without substantial loss in effectiveness which demonstrates they are over-parameterized. The contribution of this work is…

Machine Learning · Statistics 2021-03-09 Yaniv Shulman

In a previous article, a least square regression estimation procedure was proposed: first, we condiser a family of functions and study the properties of an estimator in every unidimensionnal model defined by one of these functions; we then…

Statistics Theory · Mathematics 2007-06-13 Pierre Alquier

The inversion of electromagnetic induction data to a conductivity profile is an ill-posed problem. Regularization improves the stability of the inversion and, based on Occam's razor principle, a smoothing constraint is typically used.…

Geophysics · Physics 2021-05-19 Wouter Deleersnyder , Benjamin Maveau , Thomas Hermans , David Dudal

This paper presents minimax rates for density estimation when the data dimension $d$ is allowed to grow with the number of observations $n$ rather than remaining fixed as in previous analyses. We prove a non-asymptotic lower bound which…

Statistics Theory · Mathematics 2017-08-16 Daniel J. McDonald

It is a common contention that it is an ``impossible mission'' to exactly determine the minimum sample size for the estimation of a binomial parameter with prescribed margin of error and confidence level. In this paper, we investigate such…

Statistics Theory · Mathematics 2007-08-02 Xinjia Chen

The Minimum Description Length (MDL) principle offers a formal framework for applying Occam's razor in machine learning. However, its application to neural networks such as Transformers is challenging due to the lack of a principled,…

Machine Learning · Computer Science 2026-03-04 Peter Shaw , James Cohan , Jacob Eisenstein , Kristina Toutanova
‹ Prev 1 4 5 6 7 8 10 Next ›