English
Related papers

Related papers: Cumulative Step-size Adaptation on Linear Function…

200 papers

We study decentralized optimization over networks where agents cooperatively minimize a smooth (strongly) convex sum of local losses while communicating only with immediate neighbors. Prevailing decentralized methods require either…

Optimization and Control · Mathematics 2026-05-04 Xiaokai Chen , Ilya Kuruzov , Gesualdo Scutari

Penalized linear regression is of fundamental importance in high-dimensional statistics and has been routinely used to regress a response on a high-dimensional set of predictors. In many scientific applications, there exists external…

Methodology · Statistics 2023-02-21 Sandipan Pramanik , Xianyang Zhang

A punctuated equilibrium model of biological evolution with relative fitness between different species being the fundamental driving force of evolution is introduced. Mutation is modeled as a fitness updating cellular automaton process…

Condensed Matter · Physics 2015-06-25 H. F. Chau , L. Mak , P. K. Kwok

Existing multi-strategy adaptive differential evolution (DE) commonly involves trials of multiple strategies and then rewards better-performing ones with more resources. However, the trials of an exploitative or explorative strategy may…

Neural and Evolutionary Computing · Computer Science 2021-12-03 Sheng Xin Zhang , Wing Shing Chan , Kit Sang Tang , Shao Yong Zheng

In this paper, we study numerical approximations for stochastic differential equations (SDEs) that use adaptive step sizes. In particular, we consider a general setting where decisions to reduce step sizes are allowed to depend on the…

Numerical Analysis · Mathematics 2025-12-10 James Foster , Andraž Jelinčič

Many popular adaptive gradient methods such as Adam and RMSProp rely on an exponential moving average (EMA) to normalize their stepsizes. While the EMA makes these methods highly responsive to new gradient information, recent research has…

Machine Learning · Computer Science 2021-10-13 Brett Daley , Christopher Amato

Interest in equilibrium-based sampling methods has grown with recent advances in computational hardware and Markov state modeling (MSM) methods, yet outstanding questions remain that hinder widespread adoption. Namely, how do sampling…

Biomolecules · Quantitative Biology 2018-05-15 Maxwell I. Zimmerman , Justin R. Porter , Xianqiang Sun , Roseane R. Silva , Gregory R. Bowman

Adaptive interventions, aka dynamic treatment regimens, are sequences of pre-specified decision rules that guide the provision of treatment for an individual given information about their baseline and evolving needs, including in response…

Methodology · Statistics 2024-05-02 Wenchu Pan , Daniel Almirall , Amy M. Kilbourne , Andrew Quanbeck , Lu Wang

Parameter-Efficient Fine-Tuning (PEFT) has emerged as a practical paradigm for adapting large language models (LLMs) without updating all parameters. Most existing approaches, such as LoRA and PiSSA, rely on low-rank decompositions of…

Machine Learning · Computer Science 2026-02-13 Songtao Wei , Yi Li , Bohan Zhang , Zhichun Guo , Ying Huang , Yuede Ji , Miao Yin , Guanpeng Li , Bingzhe Li

Quality gain is the expected relative improvement of the function value in a single step of a search algorithm. Quality gain analysis reveals the dependencies of the quality gain on the parameters of a search algorithm, based on which one…

Optimization and Control · Mathematics 2018-05-14 Youhei Akimoto , Anne Auger , Nikolaus Hansen

Previous studies on two-timescale stochastic approximation (SA) mainly focused on bounding mean-squared errors under diminishing stepsize schemes. In this work, we investigate {\it constant} stpesize schemes through the lens of Markov…

Systems and Control · Electrical Eng. & Systems 2025-02-25 Jeongyeol Kwon , Luke Dotson , Yudong Chen , Qiaomin Xie

The convergence behavior of Stochastic Gradient Descent (SGD) crucially depends on the stepsize configuration. When using a constant stepsize, the SGD iterates form a Markov chain, enjoying fast convergence during the initial transient…

Machine Learning · Computer Science 2024-12-17 Xiang Li , Qiaomin Xie

Pre-training a diverse set of neural network controllers in simulation has enabled robots to adapt online to damage in robot locomotion tasks. However, finding diverse, high-performing controllers requires expensive network training and…

Robotics · Computer Science 2023-09-19 Bryon Tjanaka , Matthew C. Fontaine , David H. Lee , Aniruddha Kalkar , Stefanos Nikolaidis

We develop and analyze stochastic variants of ISTA and a full backtracking FISTA algorithms [Beck and Teboulle, 2009, Scheinberg et al., 2014] for composite optimization without the assumption that stochastic gradient is an unbiased…

Optimization and Control · Mathematics 2025-02-14 Lam M. Nguyen , Katya Scheinberg , Trang H. Tran

The Canonical Correlation Analysis (CCA) family of methods is foundational in multiview learning. Regularised linear CCA methods can be seen to generalise Partial Least Squares (PLS) and be unified with a Generalized Eigenvalue Problem…

Machine Learning · Computer Science 2024-05-02 James Chapman , Lennie Wells , Ana Lawry Aguila

In this paper, we present Adaptive Computation Steps (ACS) algo-rithm, which enables end-to-end speech recognition models to dy-namically decide how many frames should be processed to predict a linguistic output. The model that applies ACS…

Audio and Speech Processing · Electrical Eng. & Systems 2018-09-27 Mohan Li , Min Liu , Masanori Hattori

Invariable step size based least-mean-square error (ISS-LMS) was considered as a very simple adaptive filtering algorithm and hence it has been widely utilized in many applications, such as adaptive channel estimation. It is well known that…

Information Theory · Computer Science 2015-01-29 Beiyi Liu , Guan Gui , Li Xu , Nobuhiro Shimoi

Computational multi-scale methods capitalize on a large time-scale separation to efficiently simulate slow dynamics over long time intervals. For stochastic systems, one often aims at resolving the statistics of the slowest dynamics. This…

Numerical Analysis · Mathematics 2021-05-14 Kristian Debrabant , Giovanni Samaey , Przemysław Zieliński

This paper proposes an improved variable step-size (VSS) algorithm for the recently introduced affine projection sign algorithm (APSA) based on the recovery of the near-end signal energy in the error signal. Simulation results demonstrate…

Other Computer Science · Computer Science 2013-11-22 Jianming Liu , Steven L. Grant

This paper presents a model based on an hybrid system to numerically simulate the climbing phase of an aircraft. This model is then used within a trajectory prediction tool. Finally, the Covariance Matrix Adaptation Evolution Strategy…

Artificial Intelligence · Computer Science 2012-12-18 Areski Hadjaz , Gaétan Marceau , Pierre Savéant , Marc Schoenauer