English
Related papers

Related papers: On the Rademacher Complexity of Linear Hypothesis …

200 papers

In this paper, based on the theory of defining sets, a class of four-weight or five-weight linear codes over Fp is constructed. The complete weight enumerators of the linear codes are determined by means of Weil sums. In some case, there is…

Information Theory · Computer Science 2020-12-10 Xina Zhang

We show that if the nearly-linear time solvers for Laplacian matrices and their generalizations can be extended to solve just slightly larger families of linear systems, then they can be used to quickly solve all systems of linear equations…

Computational Complexity · Computer Science 2017-09-26 Rasmus Kyng , Peng Zhang

In this paper, we investigate the Rademacher complexity of deep sparse neural networks, where each neuron receives a small number of inputs. We prove generalization bounds for multilayered sparse ReLU neural networks, including…

Machine Learning · Computer Science 2023-01-31 Tomer Galanti , Mengjia Xu , Liane Galanti , Tomaso Poggio

We study bilinear embedding models for the task of multi-relational link prediction and knowledge graph completion. Bilinear models belong to the most basic models for this task, they are comparably efficient to train and use, and they can…

Machine Learning · Computer Science 2017-09-15 Yanjie Wang , Rainer Gemulla , Hui Li

Agentic theorem provers often introduce intermediate lemmas, proof sketches, or subgoal decompositions before returning to tactic-level search. This can look like an expensive detour: if proving lemmas is itself hard, why should a learned…

Machine Learning · Computer Science 2026-05-11 Sho Sonoda , Shunta Akiyama , Yuya Uezato

Recently, contrastive learning has found impressive success in advancing the state of the art in solving various machine learning tasks. However, the existing generalization analysis is very limited or even not meaningful. In particular,…

Machine Learning · Computer Science 2023-03-01 Yunwen Lei , Tianbao Yang , Yiming Ying , Ding-Xuan Zhou

We introduce Transductive Local Complexity (TLC) to extend the classical Local Rademacher Complexity (LRC) to the transductive setting, incorporating substantial and novel components. Although LRC has been used to obtain sharp…

Machine Learning · Statistics 2026-02-06 Yingzhen Yang

We consider the problem of evaluation of the weight enumerator of a binary linear code. We show that the exact evaluation is hard for polynomial hierarchy. More exactly, if WE is an oracle answering the solution of the evaluation problem…

Computational Complexity · Computer Science 2007-05-23 M. N. Vyalyi

We study the fundamental problem of learning with respect to the squared loss in a convex class. The state-of-the-art sample complexity estimates in this setting rely on Rademacher complexities, which are generally difficult to control. We…

Statistics Theory · Mathematics 2025-02-24 Daniel Bartl , Shahar Mendelson

Bogdan et al. established a new criterion to determine the existence of a maximum likelihood estimator in discrete exponential families. It uses the notion of the set of uniqueness, which allows to apply the problem to the Ising model from…

Statistics Theory · Mathematics 2025-11-27 Tomasz Skalski , Tomasz Stroiński

This work introduces Bilinear Classes, a new structural framework, which permit generalization in reinforcement learning in a wide variety of settings through the use of function approximation. The framework incorporates nearly all existing…

Machine Learning · Computer Science 2021-07-13 Simon S. Du , Sham M. Kakade , Jason D. Lee , Shachar Lovett , Gaurav Mahajan , Wen Sun , Ruosong Wang

We study the sample complexity of learning ReLU neural networks from the point of view of generalization. Given norm constraints on the weight matrices, a common approach is to estimate the Rademacher complexity of the associated function…

Machine Learning · Computer Science 2024-02-06 Mark Sellke

We study two-layer belief networks of binary random variables in which the conditional probabilities Pr[childlparents] depend monotonically on weighted sums of the parents. In large networks where exact probabilistic inference is…

Machine Learning · Computer Science 2013-02-01 Michael Kearns , Lawrence Saul

Most existing literature on supervised machine learning assumes that the training dataset is drawn from an i.i.d. sample. However, many real-world problems exhibit temporal dependence and strong correlations between the marginal…

Machine Learning · Statistics 2025-06-18 Nikola Sandrić

In this report, we explore the data selection leading to a family of estimators maximizing a centrality. The family allows a nice properties leading to accurate and robust probability density function fitting according to some criteria we…

Statistics Theory · Mathematics 2024-04-10 Djemel Ziou

Linear mixed models (LMMs) are used extensively to model dependecies of observations in linear regression and are used extensively in many application areas. Parameter estimation for LMMs can be computationally prohibitive on big data.…

Machine Learning · Statistics 2019-03-08 Zilong Tan , Kimberly Roche , Xiang Zhou , Sayan Mukherjee

Transductive learning considers situations when a learner observes $m$ labelled training points and $u$ unlabelled test points with the final goal of giving correct answers for the test points. This paper introduces a new complexity measure…

Machine Learning · Statistics 2016-02-24 Ilya Tolstikhin , Nikita Zhivotovskiy , Gilles Blanchard

We first introduce a family of binary $pq^2$-periodic sequences based on the Euler quotients modulo $pq$, where $p$ and $q$ are two distinct odd primes and $p$ divides $q-1$. The minimal polynomials and linear complexities are determined…

Information Theory · Computer Science 2022-01-10 Jingwei Zhang , Shuhong Gao , Chang-An Zhao

This paper presents a RKHS, in general, of vector-valued functions intended to be used as hypothesis space for multi-task classification. It extends similar hypothesis spaces that have previously considered in the literature. Assuming this…

Machine Learning · Computer Science 2013-12-11 Cong Li , Michael Georgiopoulos , Georgios C. Anagnostopoulos

In this short note, we provide a sample complexity lower bound for learning linear predictors with respect to the squared loss. Our focus is on an agnostic setting, where no assumptions are made on the data distribution. This contrasts with…

Machine Learning · Computer Science 2021-11-23 Ohad Shamir
‹ Prev 1 3 4 5 6 7 10 Next ›