English
Related papers

Related papers: Discrimination on the Grassmann Manifold: Fundamen…

200 papers

We study the problem of finding the index of the minimum value of a vector from noisy observations. This problem is relevant in population/policy comparison, discrete maximum likelihood, and model selection. We develop an asymptotically…

Statistics Theory · Mathematics 2026-01-21 Tianyu Zhang , Hao Lee , Jing Lei

In this paper we develop a principled, probabilistic, unified approach to non-standard classification tasks, such as semi-supervised, positive-unlabelled, multi-positive-unlabelled and noisy-label learning. We train a classifier on the…

Machine Learning · Computer Science 2020-06-17 Jeppe Nørregaard , Lars Kai Hansen

Achieving fairness in text-to-image generation demands mitigating social biases without compromising visual fidelity, a challenge critical to responsible AI. Current fairness evaluation procedures for text-to-image models rely on…

Computer Vision and Pattern Recognition · Computer Science 2025-08-26 Marco N. Bochernitsan , Rodrigo C. Barros , Lucas S. Kupssinskü

Diffusion models are powerful deep generative models, but unlike classical models, they lack an explicit low-dimensional latent space that parameterizes the data manifold. This absence makes it difficult to perform manifold-aware…

Computer Vision and Pattern Recognition · Computer Science 2026-04-03 Shinnosuke Saito , Takashi Matsubara

There are a number of hypotheses underlying the existence of adversarial examples for classification problems. These include the high-dimensionality of the data, high codimension in the ambient space of the data manifolds of interest, and…

Machine Learning · Computer Science 2024-04-15 Brian Bell , Michael Geyer , David Glickenstein , Keaton Hamm , Carlos Scheidegger , Amanda Fernandez , Juston Moore

We study the properties of "generic", in the sense of the Haar measure on the corresponding Grassmann manifold, subspaces of l^N_infinity of given dimension. We prove that every "well bounded" operator on such a subspace, say E, is a…

Functional Analysis · Mathematics 2016-09-06 P. Mankiewicz , Stanislaw J. Szarek

We propose a generalization of the asymptotic equipartition property to discrete sources with an ambiguous alphabet, and prove that it holds for irreducible stationary Markov sources with an arbitrary distinguishability relation. Our…

Information Theory · Computer Science 2022-07-29 Tamás Tasnádi , Péter Vrana

Label noise in training data can significantly degrade a model's generalization performance for supervised learning tasks. Here we focus on the problem that noisy labels are primarily mislabeled samples, which tend to be concentrated near…

Machine Learning · Computer Science 2021-03-16 Hao-Chiang Shao , Hsin-Chieh Wang , Weng-Tai Su , Chia-Wen Lin

Identification capacity has been established as a relevant performance metric for various goal-/task-oriented applications, where the receiver may be interested in only a particular message that represents an event or a task. For example,…

Information Theory · Computer Science 2025-08-27 Mohammad Javad Salariseddigh , Heinz Koeppl , Holger Boche , Vahid Jamali

We consider the problem of distinguishing between two arbitrary black-box distributions defined over the domain [n], given access to $s$ samples from both. It is known that in the worst case O(n^{2/3}) samples is both necessary and…

Data Structures and Algorithms · Computer Science 2011-10-17 Eyal Even Dar , Mark Sandler

The recent success of neural network models has shone light on a rather surprising statistical phenomenon: statistical models that perfectly fit noisy data can generalize well to unseen test data. Understanding this phenomenon of…

Machine Learning · Statistics 2022-09-13 Niladri S. Chatterji , Philip M. Long , Peter L. Bartlett

For a large class of feature maps we provide a tight asymptotic characterisation of the test error associated with learning the readout layer, in the high-dimensional limit where the input dimension, hidden layer widths, and number of…

Machine Learning · Statistics 2024-06-11 Dominik Schröder , Daniil Dmitriev , Hugo Cui , Bruno Loureiro

New non-asymptotic random coding theorems (with error probability $\epsilon$ and finite block length $n$) based on Gallager parity check ensemble and Shannon random code ensemble with a fixed codeword type are established for discrete input…

Information Theory · Computer Science 2013-03-05 En-hui Yang , Jin Meng

We define thin and asymptotically scattered metric spaces as asymptotic counterparts of discrete and scattered metric spaces respectively. We characterize asymptotically scattered spaces in terms of prohibited subspaces, and classify thin…

Combinatorics · Mathematics 2012-12-04 Igor Protasov

Since Shannon proved that Gaussian distribution is the optimum for a linear channel with additive white Gaussian noise and he calculated the corresponding channel capacity, it remains the most applied distribution in optical communications…

Information Theory · Computer Science 2016-10-26 Mariia Sorokina , Stylianos Sygletos , Sergei Turitsyn

Group fairness is a central research topic in text classification, where reaching fair treatment between sensitive groups (e.g. women vs. men) remains an open challenge. This paper presents a novel method for mitigating biases in neural…

Computation and Language · Computer Science 2023-11-22 Thibaud Leteno , Antoine Gourru , Charlotte Laclau , Rémi Emonet , Christophe Gravier

A sender wishes to broadcast a message of length $n$ over an alphabet to $r$ users, where each user $i$, $1 \leq i \leq r$ should be able to receive one of $m_i$ possible messages. The broadcast channel has noise for each of the users…

Information Theory · Computer Science 2011-12-30 Amit Weinstein

The goal of this paper is to analyze an intriguing phenomenon recently discovered in deep networks, namely their instability to adversarial perturbations (Szegedy et. al., 2014). We provide a theoretical framework for analyzing the…

Machine Learning · Computer Science 2016-03-30 Alhussein Fawzi , Omar Fawzi , Pascal Frossard

We introduce a Bayesian model for inferring mixtures of subspaces of different dimensions. The key challenge in such a mixture model is specification of prior distributions over subspaces of different dimensions. We address this challenge…

Statistics Theory · Mathematics 2015-09-24 Brian St. Thomas , Lizhen Lin , Lek-Heng Lim , Sayan Mukherjee

A binary classifier capable of abstaining from making a label prediction has two goals in tension: minimizing errors, and avoiding abstaining unnecessarily often. In this work, we exactly characterize the best achievable tradeoff between…

Machine Learning · Computer Science 2016-11-30 Akshay Balsubramani
‹ Prev 1 8 9 10 Next ›