English
Related papers

Related papers: Passivity and Immersion based-modified gradient es…

200 papers

Exploding predictive AI has enabled fast yet effective evaluation and decision-making in modern chip physical design flows. State-of-the-art frameworks typically include the objective of minimizing the mean square error (MSE) between the…

Machine Learning · Computer Science 2024-02-12 Haoyu Yang , Anthony Agnesina , Haoxing Ren

Transformers empirically perform precise probabilistic reasoning in carefully constructed ``Bayesian wind tunnels'' and in large-scale language models, yet the mechanisms by which gradient-based learning creates the required internal…

Machine Learning · Statistics 2026-05-19 Naman Agarwal , Siddhartha R. Dalal , Vishal Misra

Learning in models with discrete latent variables is challenging due to high variance gradient estimators. Generally, approaches have relied on control variates to reduce the variance of the REINFORCE estimator. Recent work (Jang et al.…

Machine Learning · Computer Science 2017-11-07 George Tucker , Andriy Mnih , Chris J. Maddison , Dieterich Lawson , Jascha Sohl-Dickstein

This paper is concerned with the computing efficiency of model predictive control (MPC) problems for dynamical systems with both rate and amplitude constraints on the inputs. Instead of augmenting the decision variables of the underlying…

Optimization and Control · Mathematics 2020-03-13 Idris Kempf , Paul Goulart , Stephen Duncan

A celebrated method for Variational Inequalities (VIs) is Extragradient (EG), which can be viewed as a standard discrete-time integration scheme. With this view in mind, in this paper we show that EG may suffer from discretization bias when…

Machine Learning · Computer Science 2026-05-08 Zhankun Luo , M. Berk Sahin , Antesh Upadhyay , Behzad Sharif , Abolfazl Hashemi

This paper re-examines the problem of parameter estimation in Bayesian networks with missing values and hidden variables from the perspective of recent work in on-line learning [Kivinen & Warmuth, 1994]. We provide a unified framework for…

Machine Learning · Computer Science 2013-02-08 Eric Bauer , Daphne Koller , Yoram Singer

A novel efficient method for computing the Knowledge-Gradient policy for Continuous Parameters (KGCP) for deterministic optimization is derived. The differences with Expected Improvement (EI), a popular choice for Bayesian optimization of…

Computational Engineering, Finance, and Science · Computer Science 2016-08-17 Joachim van der Herten , Ivo Couckuyt , Dirk Deschrijver , Tom Dhaene

Gradient methods are widely used in optimization problems. In practice, while the smoothness parameter can be estimated utilizing techniques such as backtracking, estimating the strong convexity parameter remains a challenge; moreover, even…

Optimization and Control · Mathematics 2026-02-17 Xiaozhe Hu , Sara Pollock , Zhongqin Xue , Yunrong Zhu

This paper considers a practical scenario where a classical estimation method might have already been implemented on a certain platform when one tries to apply more advanced techniques such as moving horizon estimation (MHE). We are…

Systems and Control · Computer Science 2018-07-06 He Kong , Salah Sukkarieh

Dynamic state and parameter estimation methods for dynamic security assessment in power systems are becoming increasingly important for system operators. Usually, the data used for this type of applications stems from phasor measurement…

Systems and Control · Electrical Eng. & Systems 2022-09-01 Nicolai Lorenz-Meyer , René Suchantke , Johannes Schiffer

We propose an approach based on function evaluations and Bayesian inference to extract higher-order differential information of objective functions {from a given ensemble of particles}. Pointwise evaluation $\{V(x^i)\}_i$ of some potential…

Machine Learning · Statistics 2023-03-02 Claudia Schillings , Claudia Totzeck , Philipp Wacker

Recently there has been an increasing interest in primal-dual methods for model predictive control (MPC), which require minimizing the (augmented) Lagrangian at each iteration. We propose a novel first order primal-dual method, termed…

Optimization and Control · Mathematics 2020-12-21 Yue Yu , Purnanand Elango , Behçet Açikmeşe

This paper proposes an adaptive tracking strategy with mass-inertia estimation for aerial transportation problems of multi-rotor UAVs. The dynamic model of multi-rotor UAVs with disturbances is firstly developed with a linearly…

Systems and Control · Electrical Eng. & Systems 2022-09-20 Shuyang Shi , Yuzhu Li , Wei Dong

The goal of reinforcement learning (RL) is to let an agent learn an optimal control policy in an unknown environment so that future expected rewards are maximized. The model-free RL approach directly learns the policy based on data samples.…

Machine Learning · Statistics 2013-07-22 Syogo Mori , Voot Tangkaratt , Tingting Zhao , Jun Morimoto , Masashi Sugiyama

Several authors have recently developed risk-sensitive policy gradient methods that augment the standard expected cost minimization problem with a measure of variability in cost. These studies have focused on specific risk-measures, such as…

Artificial Intelligence · Computer Science 2015-06-09 Aviv Tamar , Yinlam Chow , Mohammad Ghavamzadeh , Shie Mannor

Online nonparametric estimators are gaining popularity due to their efficient computation and competitive generalization abilities. An important example includes variants of stochastic gradient descent. These algorithms often take one…

Statistics Theory · Mathematics 2025-07-08 Tianyu Zhang , Jing Lei

Models incorporating uncertain inputs, such as random forces or material parameters, have been of increasing interest in PDE-constrained optimization. In this paper, we focus on the efficient numerical minimization of a convex and smooth…

Optimization and Control · Mathematics 2021-06-18 Caroline Geiersbach , Winnifried Wollner

Trajectory optimization of a controlled dynamical system is an essential part of autonomy, however many trajectory optimization techniques are limited by the fidelity of the underlying parametric model. In the field of robotics, a lack of…

Systems and Control · Computer Science 2017-02-17 Manan Gandhi , Yunpeng Pan , Evangelos Theodorou

This paper presents an adaptive Distribution System State Estimation (DSSE) which relies on a Cloud-based IoT paradigm. The methodology is adaptive in terms of the rate of execution of the estimation process which varies depending on the…

Networking and Internet Architecture · Computer Science 2016-11-15 Paolo Attilio Pegoraro , Alessio Meloni , Luigi Atzori , Paolo Castello , Sara Sulis

In this paper, we present a multilevel Monte Carlo (MLMC) version of the Stochastic Gradient (SG) method for optimization under uncertainty, in order to tackle Optimal Control Problems (OCP) where the constraints are described in the form…

Optimization and Control · Mathematics 2019-12-30 Matthieu Martin , Fabio Nobile , Panagiotis Tsilifis