中文
相关论文

相关论文: Passivity and Immersion based-modified gradient es…

200 篇论文

Exploding predictive AI has enabled fast yet effective evaluation and decision-making in modern chip physical design flows. State-of-the-art frameworks typically include the objective of minimizing the mean square error (MSE) between the…

机器学习 · 计算机科学 2024-02-12 Haoyu Yang , Anthony Agnesina , Haoxing Ren

Transformers empirically perform precise probabilistic reasoning in carefully constructed ``Bayesian wind tunnels'' and in large-scale language models, yet the mechanisms by which gradient-based learning creates the required internal…

机器学习 · 统计学 2026-05-19 Naman Agarwal , Siddhartha R. Dalal , Vishal Misra

Learning in models with discrete latent variables is challenging due to high variance gradient estimators. Generally, approaches have relied on control variates to reduce the variance of the REINFORCE estimator. Recent work (Jang et al.…

机器学习 · 计算机科学 2017-11-07 George Tucker , Andriy Mnih , Chris J. Maddison , Dieterich Lawson , Jascha Sohl-Dickstein

This paper is concerned with the computing efficiency of model predictive control (MPC) problems for dynamical systems with both rate and amplitude constraints on the inputs. Instead of augmenting the decision variables of the underlying…

最优化与控制 · 数学 2020-03-13 Idris Kempf , Paul Goulart , Stephen Duncan

A celebrated method for Variational Inequalities (VIs) is Extragradient (EG), which can be viewed as a standard discrete-time integration scheme. With this view in mind, in this paper we show that EG may suffer from discretization bias when…

机器学习 · 计算机科学 2026-05-08 Zhankun Luo , M. Berk Sahin , Antesh Upadhyay , Behzad Sharif , Abolfazl Hashemi

This paper re-examines the problem of parameter estimation in Bayesian networks with missing values and hidden variables from the perspective of recent work in on-line learning [Kivinen & Warmuth, 1994]. We provide a unified framework for…

机器学习 · 计算机科学 2013-02-08 Eric Bauer , Daphne Koller , Yoram Singer

A novel efficient method for computing the Knowledge-Gradient policy for Continuous Parameters (KGCP) for deterministic optimization is derived. The differences with Expected Improvement (EI), a popular choice for Bayesian optimization of…

计算工程、金融与科学 · 计算机科学 2016-08-17 Joachim van der Herten , Ivo Couckuyt , Dirk Deschrijver , Tom Dhaene

Gradient methods are widely used in optimization problems. In practice, while the smoothness parameter can be estimated utilizing techniques such as backtracking, estimating the strong convexity parameter remains a challenge; moreover, even…

最优化与控制 · 数学 2026-02-17 Xiaozhe Hu , Sara Pollock , Zhongqin Xue , Yunrong Zhu

This paper considers a practical scenario where a classical estimation method might have already been implemented on a certain platform when one tries to apply more advanced techniques such as moving horizon estimation (MHE). We are…

系统与控制 · 计算机科学 2018-07-06 He Kong , Salah Sukkarieh

Dynamic state and parameter estimation methods for dynamic security assessment in power systems are becoming increasingly important for system operators. Usually, the data used for this type of applications stems from phasor measurement…

系统与控制 · 电气工程与系统科学 2022-09-01 Nicolai Lorenz-Meyer , René Suchantke , Johannes Schiffer

We propose an approach based on function evaluations and Bayesian inference to extract higher-order differential information of objective functions {from a given ensemble of particles}. Pointwise evaluation $\{V(x^i)\}_i$ of some potential…

机器学习 · 统计学 2023-03-02 Claudia Schillings , Claudia Totzeck , Philipp Wacker

Recently there has been an increasing interest in primal-dual methods for model predictive control (MPC), which require minimizing the (augmented) Lagrangian at each iteration. We propose a novel first order primal-dual method, termed…

最优化与控制 · 数学 2020-12-21 Yue Yu , Purnanand Elango , Behçet Açikmeşe

This paper proposes an adaptive tracking strategy with mass-inertia estimation for aerial transportation problems of multi-rotor UAVs. The dynamic model of multi-rotor UAVs with disturbances is firstly developed with a linearly…

系统与控制 · 电气工程与系统科学 2022-09-20 Shuyang Shi , Yuzhu Li , Wei Dong

The goal of reinforcement learning (RL) is to let an agent learn an optimal control policy in an unknown environment so that future expected rewards are maximized. The model-free RL approach directly learns the policy based on data samples.…

机器学习 · 统计学 2013-07-22 Syogo Mori , Voot Tangkaratt , Tingting Zhao , Jun Morimoto , Masashi Sugiyama

Several authors have recently developed risk-sensitive policy gradient methods that augment the standard expected cost minimization problem with a measure of variability in cost. These studies have focused on specific risk-measures, such as…

人工智能 · 计算机科学 2015-06-09 Aviv Tamar , Yinlam Chow , Mohammad Ghavamzadeh , Shie Mannor

Online nonparametric estimators are gaining popularity due to their efficient computation and competitive generalization abilities. An important example includes variants of stochastic gradient descent. These algorithms often take one…

统计理论 · 数学 2025-07-08 Tianyu Zhang , Jing Lei

Models incorporating uncertain inputs, such as random forces or material parameters, have been of increasing interest in PDE-constrained optimization. In this paper, we focus on the efficient numerical minimization of a convex and smooth…

最优化与控制 · 数学 2021-06-18 Caroline Geiersbach , Winnifried Wollner

Trajectory optimization of a controlled dynamical system is an essential part of autonomy, however many trajectory optimization techniques are limited by the fidelity of the underlying parametric model. In the field of robotics, a lack of…

系统与控制 · 计算机科学 2017-02-17 Manan Gandhi , Yunpeng Pan , Evangelos Theodorou

This paper presents an adaptive Distribution System State Estimation (DSSE) which relies on a Cloud-based IoT paradigm. The methodology is adaptive in terms of the rate of execution of the estimation process which varies depending on the…

网络与互联网体系结构 · 计算机科学 2016-11-15 Paolo Attilio Pegoraro , Alessio Meloni , Luigi Atzori , Paolo Castello , Sara Sulis

In this paper, we present a multilevel Monte Carlo (MLMC) version of the Stochastic Gradient (SG) method for optimization under uncertainty, in order to tackle Optimal Control Problems (OCP) where the constraints are described in the form…

最优化与控制 · 数学 2019-12-30 Matthieu Martin , Fabio Nobile , Panagiotis Tsilifis