English
Related papers

Related papers: Incorporating Local Step-Size Adaptivity into the …

200 papers

Gibbs sampling, as a model learning method, is known to produce the most accurate results available in a variety of domains, and is a de facto standard in these domains. Yet, it is also well known that Gibbs random walks usually have…

Machine Learning · Statistics 2018-04-20 Mark Kozdoba , Shie Mannor

Gait adaptation is an important part of gait analysis and its neuronal origin and dynamics has been studied extensively. In neurorehabilitation, it is important as it perturbs neuronal dynamics and allows patients to restore some of their…

Human-Computer Interaction · Computer Science 2021-10-19 Ines Domingos , Guang-Zhong Yang , Fani Deligianni

We suggest a simple adaptive step-size procedure, which does not require any line-search, for a general class of nonlinear optimization methods and prove convergence of a general method under mild assumptions. In particular, the goal…

Optimization and Control · Mathematics 2018-03-05 Igor Konnov

A local convergence analysis of the Gauss-Newton method for solving injective-overdetermined systems of nonlinear equations under a majorant condition is provided. The convergence as well as results on its rate are established without a…

Optimization and Control · Mathematics 2013-03-21 Max Leandro Nobre Goncalves

Penalty methods relax the incompressibility condition and uncouple velocity and pressure. Experience with them indicates that the velocity error is sensitive to the choice of penalty parameter $\epsilon$. So far, there is no effective \'a…

Numerical Analysis · Mathematics 2024-04-19 Rui Fang

In decentralized optimization, the choice of stepsize plays a critical role in algorithm performance. A common approach is to use a shared stepsize across all agents to ensure convergence. However, selecting an optimal stepsize often…

Optimization and Control · Mathematics 2026-01-07 Diyako Ghaderyan , Stefan Werner

Anytime neural networks (AnytimeNNs) are a promising solution to adaptively adjust the model complexity at runtime under various hardware resource constraints. However, the manually-designed AnytimeNNs are biased by designers' prior…

Machine Learning · Computer Science 2023-06-21 Guihong Li , Kartikeya Bhardwaj , Yuedong Yang , Radu Marculescu

A central task in many applications is reasoning about processes that change over continuous time. Continuous-Time Bayesian Networks is a general compact representation language for multi-component continuous-time processes. However, exact…

Artificial Intelligence · Computer Science 2012-06-18 Tal El-Hay , Nir Friedman , Raz Kupferman

When preparing a Gibbs sampler, some conditionals may be unfamiliar distributions without well-known variate generation routines. Rejection sampling may be used to draw from such distributions exactly; however, it can be challenging to…

Methodology · Statistics 2025-09-23 Andrew M. Raim , Kyle M. Irimata , James A. Livsey

It has been widely documented that the sampling and resampling steps in particle filters cannot be differentiated. The {\itshape reparameterisation trick} was introduced to allow the sampling step to be reformulated into a differentiable…

Machine Learning · Statistics 2022-08-10 Conor Rosato , Vincent Beraud , Paul Horridge , Thomas B. Schön , Simon Maskell

In this paper, we propose a simple, fast and easy to implement algorithm LOSSGRAD (locally optimal step-size in gradient descent), which automatically modifies the step-size in gradient descent during neural networks training. Given a…

Machine Learning · Computer Science 2019-11-26 Bartosz Wójcik , Łukasz Maziarka , Jacek Tabor

We propose an adaptive importance sampling scheme for Gaussian approximations of intractable posteriors. Optimization-based approximations like variational inference can be too inaccurate while existing Monte Carlo methods can be too slow.…

Computation · Statistics 2025-02-04 Willem van den Boom , Andrea Cremaschi , Alexandre H. Thiery

In this work, we propose an adaptive spectral element algorithm for solving nonlinear optimal control problems. The method employs orthogonal collocation at the shifted Gegenbauer-Gauss points combined with very accurate and stable…

Optimization and Control · Mathematics 2023-03-06 Kareem T. Elgindy

This work proposes an adaptive framework to solve a robust structural shape optimization problem governed by linear elasticity models that account for uncertainties in the loading and material inputs. A posteriori error estimators are…

Optimization and Control · Mathematics 2026-02-06 Oğuz Han Altıntaş , Hamdullah Yücel

Gibbs sampling methods are standard tools to perform posterior inference for mixture models. These have been broadly classified into two categories: marginal and conditional methods. While conditional samplers are more widely applicable…

Methodology · Statistics 2023-02-21 Pierpaolo De Blasi , María F. Gil-Leyva

The validity of estimation and smoothing parameter selection for the wide class of generalized additive models for location, scale and shape (GAMLSS) relies on the correct specification of a likelihood function. Deviations from such…

Methodology · Statistics 2019-11-14 William H. Aeberhard , Eva Cantoni , Giampiero Marra , Rosalba Radice

Vanilla gradient methods are often highly sensitive to the choice of stepsize, which typically requires manual tuning. Adaptive methods alleviate this issue and have therefore become widely used. Among them, AdaGrad has been particularly…

Machine Learning · Statistics 2026-02-16 Matia Bojovic , Saverio Salzo , Massimiliano Pontil

The recently introduced Matching by Tone Mapping (MTM) dissimilarity measure enables template matching under smooth non-linear distortions and also has a well-established mathematical background. MTM operates by binning the template, but…

Machine Learning · Computer Science 2020-07-27 Attila Fazekas , György Kovács

We study the fundamental optimization principles of self-attention, the defining mechanism of transformers, by analyzing the implicit bias of gradient-based optimizers in training a self-attention layer with a linear decoder in binary…

Machine Learning · Computer Science 2025-04-01 Bhavya Vasudeva , Puneesh Deora , Christos Thrampoulidis

We present a performant, general-purpose gradient-guided nested sampling algorithm, ${\tt GGNS}$, combining the state of the art in differentiable programming, Hamiltonian slice sampling, clustering, mode separation, dynamic nested…

Machine Learning · Computer Science 2023-12-08 Pablo Lemos , Nikolay Malkin , Will Handley , Yoshua Bengio , Yashar Hezaveh , Laurence Perreault-Levasseur
‹ Prev 1 4 5 6 7 8 10 Next ›