English
Related papers

Related papers: Information theoretic limits of learning a sparse …

200 papers

We propose a new approach for metric learning by framing it as learning a sparse combination of locally discriminative metrics that are inexpensive to generate from the training data. This flexible framework allows us to naturally derive…

Machine Learning · Computer Science 2019-01-25 Yuan Shi , Aurélien Bellet , Fei Sha

Fundamental limitations or performance trade-offs/limits are important properties and constraints of both control and filtering systems. Among various trade-off metrics, total information rate that characterizes the sensitivity trade-offs…

Systems and Control · Electrical Eng. & Systems 2025-03-18 Neng Wan , Dapeng Li , Naira Hovakimyan , Petros G. Voulgaris

We consider the problem of learning a target function corresponding to a single hidden layer neural network, with a quadratic activation function after the first layer, and random weights. We consider the asymptotic limit where the input…

Machine Learning · Statistics 2025-02-10 Antoine Maillard , Emanuele Troiani , Simon Martin , Florent Krzakala , Lenka Zdeborová

This work studies the properties of the maximum likelihood estimator (MLE) of a non-linear model with Gaussian errors and multidimensional parameter. The observations are collected in a two-stage experimental design and are dependent since…

Statistics Theory · Mathematics 2019-11-01 Nancy Flournoy , Caterina May , Chiara Tommasi

An alternative to extrinsic information transfer (EXIT) charts called mean squared error (MSE) charts that use a measure related to the MSE instead of mutual information is proposed. Using the relationship between mutual information and…

Information Theory · Computer Science 2007-07-13 Kapil Bhattad , Krishna Narayanan

We consider machine learning techniques to develop low-latency approximate solutions to a class of inverse problems. More precisely, we use a probabilistic approach for the problem of recovering sparse stochastic signals that are members of…

Information Theory · Computer Science 2016-09-06 Steffen Limmer , Sławomir Stańczak

We formulate meta learning using information theoretic concepts; namely, mutual information and the information bottleneck. The idea is to learn a stochastic representation or encoding of the task description, given by a training set, that…

Machine Learning · Computer Science 2021-07-06 Michalis K. Titsias , Francisco J. R. Ruiz , Sotirios Nikoloutsopoulos , Alexandre Galashov

We consider tensor factorizations using a generative model and a Bayesian approach. We compute rigorously the mutual information, the Minimal Mean Squared Error (MMSE), and unveil information-theoretic phase transitions. In addition, we…

Statistics Theory · Mathematics 2020-01-22 Thibault Lesieur , Léo Miolane , Marc Lelarge , Florent Krzakala , Lenka Zdeborová

This paper proposes a new Bayesian machine learning model that can be applied to large datasets arising in macroeconomics. Our framework sums over many simple two-component location mixtures. The transition between components is determined…

Econometrics · Economics 2023-12-05 Florian Huber

A frequentist asymptotic expansion method for error estimation is employed for a network of gravitational wave detectors to assess the amount of information that can be extracted from gravitational wave observations. Mathematically we…

General Relativity and Quantum Cosmology · Physics 2016-06-22 Rhondale Tso , Michele Zanolin

Subsampling is a computationally effective approach to extract information from massive data sets when computing resources are limited. After a subsample is taken from the full data, most available methods use an inverse probability…

Statistics Theory · Mathematics 2022-10-11 HaiYing Wang , Jae Kwang Kim

We address the fundamental limits of learning unknown parameters of any stochastic process from time-series data, and discover exact closed-form expressions for how optimal inference scales with observation length. Given a parametrized…

Machine Learning · Computer Science 2023-10-09 Paul M. Riechers

We study the evolution of conditional mutual information in generic open quantum systems, focusing on one-dimensional random circuits with interspersed local noise. Unlike in noiseless circuits, where conditional mutual information spreads…

Quantum Physics · Physics 2024-11-18 Su-un Lee , Changhun Oh , Yat Wong , Senrui Chen , Liang Jiang

Minimum mean square error (MMSE) estimation is widely used in signal processing and related fields. While it is known to be non-continuous with respect to all standard notions of stochastic convergence, it remains robust in practical…

Signal Processing · Electrical Eng. & Systems 2025-05-01 Elad Domanovitz , Anatoly Khina

Missing data imputation, where a model is trained on observed data to estimate unobserved values, is a fundamental problem in machine learning. In this paper, we rigorously formulate imputation model learning as a mean-squared error risk…

Machine Learning · Statistics 2026-05-14 Luke Shannon , Song Liu , Katarzyna Reluga

Estimating linear, mean-square continuous functionals is a pivotal challenge in statistics. In high-dimensional contexts, this estimation is often performed under the assumption of exact model sparsity, meaning that only a small number of…

Statistics Theory · Mathematics 2025-08-04 Jelena Bradic , Victor Chernozhukov , Whitney K. Newey , Yinchu Zhu

This paper introduces a novel framework and corresponding methods for sampling and reconstruction of sparse signals in shift-invariant (SI) spaces. We reinterpret the random demodulator, a system that acquires sparse bandlimited signals, as…

Signal Processing · Electrical Eng. & Systems 2022-01-24 Tin Vlašić , Damir Seršić

Nonparametric regression problems with qualitative constraints such as monotonicity or convexity are ubiquitous in applications. For example, in predicting the yield of a factory in terms of the number of labor hours, the monotonicity of…

Statistics Theory · Mathematics 2023-11-21 Soham Mallick , Siddhaarth Sarkar , Arun Kumar Kuchibhotla

Ising models describe the joint probability distribution of a vector of binary feature variables. Typically, not all the variables interact with each other and one is interested in learning the presumably sparse network structure of the…

Machine Learning · Computer Science 2019-07-09 Frank Nussbaum , Joachim Giesen

Minimizing the Mean Squared Error (MSE) is a key objective in machine learning and is commonly used for imputing missing values. While this approach provides accurate point estimates, it introduces systematic biases in downstream analyses.…

Machine Learning · Statistics 2026-05-06 Stef van Buuren