English
Related papers

Related papers: The Bernstein-Orlicz norm and deviation inequaliti…

200 papers

Generalized gamma distributions arise as limits in many settings involving random graphs, walks, trees, and branching processes. Pek\"oz, R\"ollin, and Ross (2016, arXiv:1309.4183 [math.PR]) exploited characterizing distributional fixed…

Probability · Mathematics 2022-08-08 Tobias Johnson , Erol Peköz

We derive generalization and excess risk bounds for neural nets using a family of complexity measures based on a multilevel relative entropy. The bounds are obtained by introducing the notion of generated hierarchical coverings of neural…

Machine Learning · Computer Science 2019-06-27 Amir R. Asadi , Emmanuel Abbe

Renyi's "thinning" operation on a discrete random variable is a natural discrete analog of the scaling operation for continuous random variables. The properties of thinning are investigated in an information-theoretic context, especially in…

Information Theory · Computer Science 2010-08-17 Peter Harremoes , Oliver Johnson , Ioannis Kontoyiannis

We prove uniform estimates for the expected value of averages of order statistics of bivariate functions in terms of their largest values by a direct analysis. As an application, uniform estimates for the expected value of averages of order…

Probability · Mathematics 2018-10-03 Richard Lechner , Markus Passenbrunner , Joscha Prochno

Time series regression models are commonly used in time series analysis. However, in modern real-world applications, serially correlated data with an ultra-high dimension and fat tails are prevalent. This presents a challenge in developing…

Statistics Theory · Mathematics 2023-04-21 Linbo Liu , Danna Zhang

This paper considers the entropy of the sum of (possibly dependent and non-identically distributed) Bernoulli random variables. Upper bounds on the error that follows from an approximation of this entropy by the entropy of a Poisson random…

Information Theory · Computer Science 2016-11-17 Igal Sason

The recent success of neural network models has shone light on a rather surprising statistical phenomenon: statistical models that perfectly fit noisy data can generalize well to unseen test data. Understanding this phenomenon of…

Machine Learning · Statistics 2022-09-13 Niladri S. Chatterji , Philip M. Long , Peter L. Bartlett

Measuring dependence between random variables is a fundamental problem in Statistics, with applications across diverse fields. While classical measures such as Pearson's correlation have been widely used for over a century, they have…

Statistics Theory · Mathematics 2025-10-08 Marta Catalano , Hugo Lavenant

We are concerned with obtaining novel concentration inequalities for the missing mass, i.e. the total probability mass of the outcomes not observed in the sample. We not only derive - for the first time - distribution-free Bernstein-like…

Machine Learning · Statistics 2015-06-22 Bahman Yari Saeed Khanloo , Gholamreza Haffari

Model selection is often performed by empirical risk minimization. The quality of selection in a given situation can be assessed by risk bounds, which require assumptions both on the margin and the tails of the losses used. Starting with…

Statistics Theory · Mathematics 2008-12-18 Charles Mitchell , Sara van de Geer

We consider a generalization of the classic linear regression problem to the case when the loss is an Orlicz norm. An Orlicz norm is parameterized by a non-negative convex function $G:\mathbb{R}_+\rightarrow\mathbb{R}_+$ with $G(0)=0$: the…

Data Structures and Algorithms · Computer Science 2018-06-19 Alexandr Andoni , Chengyu Lin , Ying Sheng , Peilin Zhong , Ruiqi Zhong

For obtaining causal inferences that are objective, and therefore have the best chance of revealing scientific truths, carefully designed and executed randomized experiments are generally considered to be the gold standard. Observational…

Applications · Statistics 2008-11-12 Donald B. Rubin

Markov chains are a natural and well understood tool for describing one-dimensional patterns in time or space. We show how to infer $k$-th order Markov chains, for arbitrary $k$, from finite data by applying Bayesian methods to both…

Statistics Theory · Mathematics 2009-11-13 Christopher C. Strelioff , James P. Crutchfield , Alfred W. Hubler

While effective concentration inequalities for suprema of empirical processes exist under boundedness or strict tail assumptions, no comparable results have been available under considerably weaker assumptions. In this paper, we derive…

Probability · Mathematics 2014-10-23 Johannes Lederer , Sara van de Geer

Random-cluster measures on infinite regular trees are studied in conjunction with a general type of `boundary condition', namely an equivalence relation on the set of infinite paths of the tree. The uniqueness and non-uniqueness of…

Probability · Mathematics 2007-05-23 Geoffrey Grimmett , Svante Janson

This paper develops techniques to study the number of descents in random permutations via martingales. We relax an assumption in the Berry-Esseen theorem of Bolthausen (1982) to extend the theorem's scope to martingale differences of…

Probability · Mathematics 2021-03-16 Alperen Y. Özdemir

We derive some key extremal features for $k$th order Markov chains that can be used to understand how the process moves between an extreme state and the body of the process. The chains are studied given that there is an exceedance of a…

Statistics Theory · Mathematics 2023-01-27 Ioannis Papastathopoulos , Adrian Casey , Jonathan A. Tawn

Optimal transport (OT) is a versatile framework for comparing probability measures, with many applications to statistics, machine learning, and applied mathematics. However, OT distances suffer from computational and statistical scalability…

Statistics Theory · Mathematics 2022-06-08 Ziv Goldfeld , Kengo Kato , Gabriel Rioux , Ritwik Sadhu

We consider the stochastic integrals of multivariate point processes and study their concentration phenomena. In particular, we obtain a Bernstein type of concentration inequality through Dol\'eans-Dade exponential formula and a uniform…

Probability · Mathematics 2017-03-24 Hanchao Wang , Zhengyan Lin , Zhonggen Su

We prove a moderate deviation principle for subgraph count statistics of Erdos-Renyi random graphs. This is equivalent in showing a moderate deviation principle for the trace of a power of a Bernoulli random matrix. It is done via an…

Probability · Mathematics 2010-03-31 Hanna Döring , Peter Eichelsbacher