English
Related papers

Related papers: Verifying Good Regulator Conditions for Hypergraph…

200 papers

Pairwise Fisher graphs capture local covariance information, but they cannot distinguish an irreducible multi-observable radiation pattern from a collection of ordinary pairwise correlations. We show that this missing structure is naturally…

High Energy Physics - Phenomenology · Physics 2026-05-08 Aritra Bal , Markus Klute , Benedikt Maier , Michael Spannowsky

Gradient-flow (GF) viewpoints unify and illuminate optimization algorithms, yet most GF analyses focus on unconstrained settings. We develop a geometry-respecting framework for constrained problems by (i) reparameterizing feasible sets with…

Optimization and Control · Mathematics 2025-08-29 Valentin Leplat

Adaptive gradient methods, such as AdaGrad, are among the most successful optimization algorithms for neural network training. While these methods are known to achieve better dimensional dependence than stochastic gradient descent (SGD) for…

Optimization and Control · Mathematics 2025-06-09 Ruichen Jiang , Devyani Maladkar , Aryan Mokhtari

Learning in neural networks is often framed as a problem in which targeted error signals are directly propagated to parameters and used to produce updates that induce more optimal network behaviour. Backpropagation of error (BP) is an…

Neural and Evolutionary Computing · Computer Science 2023-01-30 Nasir Ahmad , Ellen Schrader , Marcel van Gerven

Motivated by a wide variety of applications, ranging from stochastic optimization to dimension reduction through variable selection, the problem of estimating gradients accurately is of crucial importance in statistics and learning theory.…

Machine Learning · Computer Science 2020-06-29 Guillaume Ausset , Stephan Clémençon , François Portier

We undertake Bayesian learning of the high-dimensional functional relationship between a system parameter vector and an observable, that is in general tensor-valued. The ultimate aim is Bayesian inverse prediction of the system parameters,…

Methodology · Statistics 2018-04-17 Kangrui Wang , Dalia Chakrabarty

We study the problem of learning a directed acyclic graph from data generated according to an additive, non-linear structural equation model with Gaussian noise. We express each non-linear function through a basis expansion, and derive a…

Methodology · Statistics 2025-11-27 Xiaozhu Zhang , Nir Keret , Ali Shojaie , Armeen Taeb

The need for large amounts of training data in modern machine learning is one of the biggest challenges of the field. Compared to the brain, current artificial algorithms are much less capable of learning invariance transformations and…

Neural and Evolutionary Computing · Computer Science 2023-07-25 Aleksandar Vučković , Benedikt Stock , Alexander V. Hopp , Mathias Winkel , Helmut Linde

In a recent paper [WW23] we studied the transport of oscillations in solutions to linear and some semilinear second-order hyperbolic boundary problems along rays that graze a convex obstacle to any order. We showed that high frequency exact…

Analysis of PDEs · Mathematics 2026-02-25 Jian Wang , Mark Williams

Learning rules -- prescriptions for updating model parameters to improve performance -- are typically assumed rather than derived. Why do some learning rules work better than others, and under what assumptions can a given rule be considered…

Machine Learning · Computer Science 2025-11-03 John J. Vastola , Samuel J. Gershman , Kanaka Rajan

Invariants and conservation laws convey critical information about the underlying dynamics of a system, yet it is generally infeasible to find them from large-scale data without any prior knowledge or human insight. We propose ConservNet to…

Machine Learning · Computer Science 2021-07-01 Seungwoong Ha , Hawoong Jeong

We consider the problem of learning the structure of ferromagnetic Ising models Markov on sparse Erdos-Renyi random graph. We propose simple local algorithms and analyze their performance in the regime of correlation decay. We prove that an…

Statistics Theory · Mathematics 2015-03-17 Animashree Anandkumar , Vincent Tan , Alan Willsky

In this paper, we consider supervised learning problems such as logistic regression and study the stochastic gradient method with averaging, in the usual stochastic approximation setting where observations are used only once. We show that…

Statistics Theory · Mathematics 2014-03-18 Francis Bach

Declines in cost and concerns about the environmental impact of traditional generation have boosted the penetration of renewables and non-conventional distributed energy resources into the power grid. The intermittent availability of these…

Systems and Control · Electrical Eng. & Systems 2022-03-10 Priyank Srivastava , Patricia Hidalgo-Gonzalez , Jorge Cortes

The central question of this paper is: how do algebraic invariants of edge ideals change under natural graph operations? We study this question through the lens of suspensions. The (full) suspension of a graph is obtained by adjoining a new…

Commutative Algebra · Mathematics 2026-03-09 Selvi Kara , Dalena Vien

Recently, several works have shown that natural modifications of the classical conditional gradient method (aka Frank-Wolfe algorithm) for constrained convex optimization, provably converge with a linear rate when: i) the feasible set is a…

Optimization and Control · Mathematics 2016-05-23 Dan Garber , Ofer Meshi

We investigate sufficient conditions under which cubic gravity is healthy and viable at the perturbation level. We perform a detailed analysis of the scalar and tensor perturbations. We impose the requirement that the two scalar potentials,…

General Relativity and Quantum Cosmology · Physics 2024-02-19 Petros Asimakis , Spyros Basilakos , Emmanuel N. Saridakis

It is a central challenge in deep learning to understand how neural networks learn representations. A leading approach is the Neural Feature Ansatz (NFA) (Radhakrishnan et al. 2024), a conjectured mechanism for how feature learning occurs.…

Machine Learning · Computer Science 2025-09-08 Enric Boix-Adsera , Neil Mallinar , James B. Simon , Mikhail Belkin

Ample empirical evidence in deep neural network training suggests that a variety of optimizers tend to find nearly global optima. In this article, we adopt the reversed perspective that convergence to an arbitrary point is assumed rather…

Machine Learning · Computer Science 2025-10-13 Jerome Bolte , Quoc-Tung Le , Edouard Pauwels

The method to design exponentially stable adaptive observers is proposed for linear time-invariant systems parameterized by unknown physical parameters. Unlike existing adaptive solutions, the system state-space matrices A, B are not…

Systems and Control · Electrical Eng. & Systems 2023-08-22 Anton Glushchenko , Konstantin Lastochkin