English
Related papers

Related papers: Universal compression of Gaussian sources with unk…

200 papers

Communication is one of the key bottlenecks in the distributed training of large-scale machine learning models, and lossy compression of exchanged information, such as stochastic gradients or models, is one of the most effective instruments…

Machine Learning · Computer Science 2022-06-22 Egor Shulgin , Peter Richtárik

We show a statistical version of Taylor's theorem and apply this result to non-parametric density estimation from truncated samples, which is a classical challenge in Statistics \cite{woodroofe1985estimating, stute1993almost}. The…

Statistics Theory · Mathematics 2021-07-01 Constantinos Daskalakis , Vasilis Kontonis , Christos Tzamos , Manolis Zampetakis

We discuss estimating the probability that the sum of nonnegative independent and identically distributed random variables falls below a given threshold, i.e., $\mathbb{P}(\sum_{i=1}^{N}{X_i} \leq \gamma)$, via importance sampling (IS). We…

Computation · Statistics 2021-10-04 Nadhir Ben Rached , Abdul-Lateef Haji-Ali , Gerardo Rubino , Raul Tempone

In this paper, we propose a data based transformation for infinite-dimensional Gaussian processes and derive its limit theorem. For a classification problem, this transformation induces complete separation among the associated Gaussian…

Statistics Theory · Mathematics 2022-03-25 Juan A. Cuesta-Albertos , Subhajit Dutta

We introduce a novel approach based on stochastic optimization to find the optimal sampling distribution for the data-driven stability analysis of switched linear systems. Our goal is to address limitations of existing approaches, in…

Optimization and Control · Mathematics 2025-09-01 Alexis Vuille , Guillaume O. Berger , Raphaël M. Jungers

This paper establishes the optimal sub-Gaussian variance proxy for truncated Gaussian and truncated exponential random variables. The proofs rely on first characterizing the optimal variance proxy as the unique solution to a set of two…

Statistics Theory · Mathematics 2024-11-27 Mathias Barreto , Olivier Marchal , Julyan Arbel

Given $n$ samples from a population of individuals belonging to different types with unknown proportions, how do we estimate the probability of discovering a new type at the $(n+1)$-th draw? This is a classical problem in statistics,…

Statistics Theory · Mathematics 2018-06-27 Fadhel Ayed , Marco Battiston , Federico Camerlenghi , Stefano Favaro

Source confusion has been a long-standing problem in the astronomical history. In the previous formulation, sources are assumed to be distributed homogeneously on the sky. This fundamental assumption is not realistic in many applications.…

Astrophysics · Physics 2009-11-10 Tsutomu T. Takeuchi , Takako T. Ishii

Perturbation theory makes it possible to calculate the probability distribution function (PDF) of the large scale density field in the small variance limit. For top hat smoothing and scale-free Gaussian initial fluctuations, the result…

Astrophysics · Physics 2015-06-24 S. Colombi , F. Bernardeau , F. R. Bouchet , L. Hernquist

We develop a new approach for distributed distance computation in planar graphs that is based on a variant of the metric compression problem recently introduced by Abboud et al. [SODA'18]. One of our key technical contributions is in…

Data Structures and Algorithms · Computer Science 2019-12-30 Jason Li , Merav Parter

We study universal decoding over unknown discrete additive channels determined by a finite-state (unifilar) random process. Aiming at low-complexity decoders, we study variants of noise-guessing decoders that use estimators for the…

Information Theory · Computer Science 2025-07-24 Henrique K. Miyamoto , Sheng Yang

Based on the canonical correlation analysis we derive series representations of the probability density function (PDF) and the cumulative distribution function (CDF) of the information density of arbitrary Gaussian random vectors as well as…

Information Theory · Computer Science 2022-07-13 Jonathan Huffmann , Martin Mittelbach

We propose and study a family of universal sequential probability assignments on individual sequences, based on the incremental parsing procedure of the Lempel-Ziv (LZ78) compression algorithm. We show that the normalized log loss under any…

Information Theory · Computer Science 2025-12-15 Naomi Sagan , Tsachy Weissman

The evolution of probability distribution functions (PDFs) of continuous density, velocity and velocity derivatives ( deformation tensor) fields in the theory of cosmological gravitational instability are considered. We show that in the…

Astrophysics · Physics 2007-05-23 Lev Kofman

The universality of the directed polymer model and the analogous KPZ equation is supported by numerical simulations using non-Gaussian random probability distributions in two, three and four dimensions. It is shown that although in the…

Disordered Systems and Neural Networks · Physics 2016-08-31 Ehud Perlsman , Shlomo Havlin

Although regular expressions do not correspond univocally to regular languages, it is still worthwhile to study their properties and algorithms. For the average case analysis one often relies on the uniform random generation using a…

Formal Languages and Automata Theory · Computer Science 2021-03-25 Sabine Broda , António Machiavelo , Nelma Moreira , Rogério Reis

We consider the problem of hypothesis testing for discrete distributions. In the standard model, where we have sample access to an underlying distribution $p$, extensive research has established optimal bounds for uniformity testing,…

Machine Learning · Computer Science 2024-12-03 Maryam Aliakbarpour , Piotr Indyk , Ronitt Rubinfeld , Sandeep Silwal

We study the behavior of the posterior distribution in high-dimensional Bayesian Gaussian linear regression models having $p\gg n$, with $p$ the number of predictors and $n$ the sample size. Our focus is on obtaining quantitative finite…

Statistics Theory · Mathematics 2014-01-06 Nate Strawn , Artin Armagan , Rayan Saab , Lawrence Carin , David Dunson

It is well known that, under standard regularity conditions, the maximum likelihood estimator (MLE) satisfies a central limit theorem and converges in distribution to a Gaussian random variable as the sample size grows. This paper…

Information Theory · Computer Science 2026-05-26 Leighton P. Barnes , Alex Dytso

Recent results in compressed sensing showed that the optimal subsampling strategy should take into account the sparsity pattern of the signal at hand. This oracle-like knowledge, even though desirable, nevertheless remains elusive in most…

Information Theory · Computer Science 2023-06-28 Simon Ruetz