English
Related papers

Related papers: An Optimal Sauer Lemma Over $k$-ary Alphabets

200 papers

The multi-index model with sparse dimension reduction matrix is a popular approach to circumvent the curse of dimensionality in a high-dimensional regression setting. Building on the single-index analysis by Alquier, P. & Biau, G. (Journal…

Statistics Theory · Mathematics 2026-03-31 Maximilian F. Steffen

PAC-Bayesian is an analysis framework where the training error can be expressed as the weighted average of the hypotheses in the posterior distribution whilst incorporating the prior knowledge. In addition to being a pure generalization…

Machine Learning · Computer Science 2022-02-07 Wei Huang , Chunrui Liu , Yilan Chen , Tianyu Liu , Richard Yi Da Xu

In [1], we introduced a family of combinatorial designs, which we call "alphabet reduction pairs of arrays", ARPAs for short. These designs depend on three integer parameters $q, p \leq q, k\leq p$: $q$ is the size of the symbol set $\{0, 1…

Combinatorics · Mathematics 2024-06-18 Jean-François Culus , Sophie Toulouse

$ \newcommand{\eps}{\varepsilon} $In learning theory, the VC dimension of a concept class $C$ is the most common way to measure its "richness." In the PAC model $$ \Theta\Big(\frac{d}{\eps} + \frac{\log(1/\delta)}{\eps}\Big) $$ examples are…

Quantum Physics · Physics 2017-06-08 Srinivasan Arunachalam , Ronald de Wolf

We study the packing dimension of unions of subsets of $k$-planes in $\mathbb{R}^n$ using tools from algorithmic information theory, obtaining an analog of a result of H\'era and a mild generalization of a recent result of Fraser. Along the…

Classical Analysis and ODEs · Mathematics 2025-08-26 Jacob B. Fiedler

We prove new bounds on the dimensions of distance sets and pinned distance sets of planar sets. Among other results, we show that if $A\subset\mathbb{R}^2$ is a Borel set of Hausdorff dimension $s>1$, then its distance set has Hausdorff…

Classical Analysis and ODEs · Mathematics 2019-12-17 Tamás Keleti , Pablo Shmerkin

Bayesian classification labels observations based on given prior information, namely class-a priori and class-conditional probabilities. Bayes' risk is the minimum expected classification cost that is achieved by the Bayes' test, the…

Computer Vision and Pattern Recognition · Computer Science 2023-03-07 Frank Nielsen

Error-correcting codes resilient to synchronization errors such as insertions and deletions are known as insdel codes. Due to their important applications in DNA storage and computational biology, insdel codes have recently become a focal…

Combinatorics · Mathematics 2024-08-21 Xiangliang Kong , Itzhak Tamo , Hengjia Wei

For any family of measurable sets in a probability space, we show that either (i) the family has infinite Vapnik-Chervonenkis (VC) dimension or (ii) for every epsilon > 0 there is a finite partition pi such the pi-boundary of each set has…

Probability · Mathematics 2010-10-22 Terrence M. Adams , Andrew B. Nobel

We apply the PAC-Bayes theory to the setting of learning-to-optimize. To the best of our knowledge, we present the first framework to learn optimization algorithms with provable generalization guarantees (PAC-bounds) and explicit trade-off…

Machine Learning · Computer Science 2023-02-16 Michael Sucker , Peter Ochs

Set packing is a fundamental problem that generalises some well-known combinatorial optimization problems and knows a lot of applications. It is equivalent to hypergraph matching and it is strongly related to the maximum independent set…

Combinatorics · Mathematics 2015-07-28 Tim Oosterwijk

The main objective of this paper is to answer the questions posed by Robinson and Sadowski [21, p. 505, Comm. Math. Phys., 2010]{[RS3]} for the Navier-Stokes equations. Firstly, we prove that the upper box dimension of the potential…

Analysis of PDEs · Mathematics 2022-08-16 Yanqing Wang , Gang Wu

We present efficient counting and sampling algorithms for random $k$-SAT when the clause density satisfies $\alpha \le \frac{2^k}{\mathrm{poly}(k)}.$ In particular, the exponential term $2^k$ matches the satisfiability threshold…

Data Structures and Algorithms · Computer Science 2024-11-06 Zongchen Chen , Aditya Lonkar , Chunyang Wang , Kuan Yang , Yitong Yin

List-decoding and list-recovery are important generalizations of unique decoding that received considerable attention over the years. However, the optimal trade-off among list-decoding (resp. list-recovery) radius, list size, and the code…

Information Theory · Computer Science 2021-12-13 Eitan Goldberg , Chong Shangguan , Itzhak Tamo

A fundamental problem in adversarial machine learning is to quantify how much training data is needed in the presence of evasion attacks. In this paper we address this issue within the framework of PAC learning, focusing on the class of…

Machine Learning · Computer Science 2022-05-13 Pascale Gourdeau , Varun Kanade , Marta Kwiatkowska , James Worrell

A matrix $M: A \times X \rightarrow \{-1,1\}$ corresponds to the following learning problem: An unknown element $x \in X$ is chosen uniformly at random. A learner tries to learn $x$ from a stream of samples, $(a_1, b_1), (a_2, b_2) \ldots$,…

Machine Learning · Computer Science 2017-08-10 Sumegha Garg , Ran Raz , Avishay Tal

We study a model of machine teaching where the teacher mapping is constructed from a size function on both concepts and examples. The main question in machine teaching is the minimum number of examples needed for any concept, the so-called…

Combinatorics · Mathematics 2024-02-12 Brigt Håvardstun , Jan Kratochvíl , Joakim Sunde , Jan Arne Telle

We show that the class of strongly connected graphical models with treewidth at most k can be properly efficiently PAC-learnt with respect to the Kullback-Leibler Divergence. Previous approaches to this problem, such as those of Chow ([1]),…

Machine Learning · Computer Science 2012-07-19 Mukund Narasimhan , Jeff A. Bilmes

We present an information-theoretic framework for bounding the number of labeled samples needed to train a classifier in a parametric Bayesian setting. We derive bounds on the average $L_p$ distance between the learned classifier and the…

Information Theory · Computer Science 2017-11-20 Matthew Nokleby , Ahmad Beirami , Robert Calderbank

In this article, we consider a dominant rational self-map $f:X \dashrightarrow X$ of a normal projective variety defined over a number field. We study the arithmetic degree $\alpha_k(f)$ for $f$ and $\alpha_k(f,V)$ of a subvariety $V$,…

Number Theory · Mathematics 2025-04-15 Jiarui Song