English
Related papers

Related papers: Invariance Properties of the Natural Gradient in O…

200 papers

The idea that the brain functions so as to minimize certain costs pervades theoretical neuroscience. Since a cost function by itself does not predict how the brain finds its minima, additional assumptions about the optimization method need…

Neurons and Cognition · Quantitative Biology 2018-12-24 Simone Carlo Surace , Jean-Pascal Pfister , Wulfram Gerstner , Johanni Brea

A metric-field approach to gravitation is presented. It is based on an idea of dependency of space-time properties on measuring instruments. Some bimetric equations that realize this idea are considered. They were tested by the binary…

General Relativity and Quantum Cosmology · Physics 2009-11-07 L. V. Verozub

This paper proposes a novel parameter selection strategy for kernel-based gradient descent (KGD) algorithms, integrating bias-variance analysis with the splitting method. We introduce the concept of empirical effective dimension to quantify…

Machine Learning · Statistics 2026-03-05 Xiaotong Liu , Yunwen Lei , Xiangyu Chang , Shao-Bo Lin

In this paper, the Riemannian gradient algorithm and the natural gradient algorithm are applied to solve descent direction problems on the manifold of positive definite Hermitian matrices, where the geodesic distance is considered as the…

Optimization and Control · Mathematics 2021-06-01 Xiaomin Duan , Huafei Sun , Linyu Peng

This paper investigates the geometry of a completely integrable gradient system defined on the three parameter bivariate beta statistical manifold of the first kind. We prove that the associated vector field is Hamiltonian and admits a Lax…

Differential Geometry · Mathematics 2025-08-07 Prosper Rosaire Mama Assandje , Joseph Dongho , Thomas Bouetou Bouetou

Learning rules -- prescriptions for updating model parameters to improve performance -- are typically assumed rather than derived. Why do some learning rules work better than others, and under what assumptions can a given rule be considered…

Machine Learning · Computer Science 2025-11-03 John J. Vastola , Samuel J. Gershman , Kanaka Rajan

Optimising probabilistic models is a well-studied field in statistics. However, its connection with the training of generative models remains largely under-explored. In this paper, we show that the evolution of time-varying generative…

Machine Learning · Statistics 2025-06-16 Song Liu , Leyang Wang , Yakun Wang

The paper addresses the problem of learning a regression model parameterized by a fixed-rank positive semidefinite matrix. The focus is on the nonlinear nature of the search space and on scalability to high-dimensional problems. The…

Machine Learning · Computer Science 2011-02-01 Gilles Meyer , Silvere Bonnabel , Rodolphe Sepulchre

In this work, we propose an optimization algorithm which we call norm-adapted gradient descent. This algorithm is similar to other gradient-based optimization algorithms like Adam or Adagrad in that it adapts the learning rate of stochastic…

Machine Learning · Computer Science 2020-10-14 David Sprunger

Gradient descent can be surprisingly good at optimizing deep neural networks without overfitting and without explicit regularization. We find that the discrete steps of gradient descent implicitly regularize models by penalizing gradient…

Machine Learning · Computer Science 2022-07-20 David G. T. Barrett , Benoit Dherin

The performance of a feedforward controller is primarily determined by the extent to which it can capture the relevant dynamics of a system. The aim of this paper is to develop an input-output linear parameter-varying (LPV) feedforward…

Systems and Control · Electrical Eng. & Systems 2023-09-25 Johan Kon , Jeroen van de Wijdeven , Dennis Bruijnen , Roland Tóth , Marcel Heertjes , Tom Oomen

Learning the parameters of a (potentially partially observable) random field model is intractable in general. Instead of focussing on a single optimal parameter value we propose to treat parameters as dynamical quantities. We introduce an…

Machine Learning · Computer Science 2012-05-14 Max Welling

We propose efficient numerical schemes for implementing the natural gradient descent (NGD) for a broad range of metric spaces with applications to PDE-based optimization problems. Our technique represents the natural gradient direction as a…

Optimization and Control · Mathematics 2023-01-12 Levon Nurbekyan , Wanzhou Lei , Yunan Yang

We prove that the standard gradient flow in parameter space that underlies many training algorithms in deep learning can be continuously deformed into an adapted gradient flow which yields (constrained) Euclidean gradient flow in output…

Machine Learning · Computer Science 2026-02-02 Thomas Chen , Patrícia Muñoz Ewald

We evaluate natural gradient, an algorithm originally proposed in Amari (1997), for learning deep models. The contributions of this paper are as follows. We show the connection between natural gradient and three other recently proposed…

Machine Learning · Computer Science 2014-02-18 Razvan Pascanu , Yoshua Bengio

The Bayesian learning rule is a natural-gradient variational inference method, which not only contains many existing learning algorithms as special cases but also enables the design of new algorithms. Unfortunately, when variational…

Machine Learning · Statistics 2020-10-27 Wu Lin , Mark Schmidt , Mohammad Emtiyaz Khan

The gradient flow is the evolution of fields and physical quantities along a dimensionful parameter~$t$, the flow time. We give a simple argument that relates this gradient flow and the Wilsonian renormalization group (RG) flow. We then…

High Energy Physics - Theory · Physics 2021-07-09 Hiroki Makino , Okuto Morikawa , Hiroshi Suzuki

Estimating hyperparameters has been a long-standing problem in machine learning. We consider the case where the task at hand is modeled as the solution to an optimization problem. Here the exact gradient with respect to the hyperparameters…

Optimization and Control · Mathematics 2023-11-16 Matthias J. Ehrhardt , Lindon Roberts

Gradient matching with Gaussian processes is a promising tool for learning parameters of ordinary differential equations (ODE's). The essence of gradient matching is to model the prior over state variables as a Gaussian process which…

Machine Learning · Statistics 2016-10-25 Nico S. Gorbach , Stefan Bauer , Joachim M. Buhmann

Recently, Hammond and Sheffield introduced a model of correlated random walks that scale to fractional Brownian motions with long-range dependence. In this paper, we consider a natural generalization of this model to dimension $d\geq 2$. We…

Probability · Mathematics 2015-04-21 Hermine Biermé , Olivier Durieu , Yizao Wang
‹ Prev 1 4 5 6 7 8 10 Next ›