English
Related papers

Related papers: Return Capping: Sample-Efficient CVaR Policy Gradi…

200 papers

Transfer learning can greatly speed up reinforcement learning for a new task by leveraging policies of relevant tasks. Existing works of policy reuse either focus on only selecting a single best source policy for transfer without…

Artificial Intelligence · Computer Science 2019-03-11 Siyuan Li , Fangda Gu , Guangxiang Zhu , Chongjie Zhang

We consider a general optimal control problem in the setting of gradient flows. Two approximations of the problem are presented, both relying on the variational reformulation of gradient-flow dynamics via the Weighted-Energy-Dissipation…

Optimization and Control · Mathematics 2024-03-25 Takeshi Fukao , Ulisse Stefanelli , Riccardo Voso

This article considers the challenge of accommodating outlier measurements in state estimation. The Risk-Averse Performance-Specified (RAPS) state estimation approach addresses outliers as a measurement selection Bayesian risk minimization…

Systems and Control · Electrical Eng. & Systems 2025-05-13 Wang Hu , Zeyi Jiang , Hamed Mohsenian-Rad , Jay A. Farrell

We study the problem of optimal state-feedback tracking control for unknown discrete-time deterministic systems with input constraints. To handle input constraints, state-of-art methods utilize a certain nonquadratic stage cost function,…

Systems and Control · Electrical Eng. & Systems 2020-12-09 Alexandros Tanzanakis , John Lygeros

Model-free deep-reinforcement-based learning algorithms have been applied to a range of COPs~\cite{bello2016neural}~\cite{kool2018attention}~\cite{nazari2018reinforcement}. However, these approaches suffer from two key challenges when…

Machine Learning · Computer Science 2022-06-01 Nasrin Sultana , Jeffrey Chan , Tabinda Sarwar , A. K. Qin

We consider a liquidation problem in which a risk-averse trader tries to liquidate a fixed quantity of an asset in the presence of market impact and random price fluctuations. The trader encounters a trade-off between the transaction costs…

Trading and Market Microstructure · Quantitative Finance 2022-01-31 Seungki Min , Ciamac C. Moallemi , Costis Maglaras

Reinforcement Learning (RL) agents in the real world must satisfy safety constraints in addition to maximizing a reward objective. Model-based RL algorithms hold promise for reducing unsafe real-world actions: they may synthesize policies…

Machine Learning · Computer Science 2021-12-16 Yecheng Jason Ma , Andrew Shen , Osbert Bastani , Dinesh Jayaraman

With the pervasiveness of Stochastic Shortest-Path (SSP) problems in high-risk industries, such as last-mile autonomous delivery and supply chain management, robust planning algorithms are crucial for ensuring successful task completion…

Artificial Intelligence · Computer Science 2024-08-19 Clinton Enwerem , Erfaun Noorani , John S. Baras , Brian M. Sadler

The capacitated arc routing problem (CARP) is a challenging combinatorial optimisation problem abstracted from many real-world applications, such as waste collection, road gritting and mail delivery. However, few studies considered dynamic…

Neural and Evolutionary Computing · Computer Science 2022-02-23 Hao Tong , Leandro L. Minku , Stefan Menzel , Bernhard Sendhoff , Xin Yao

For many real-world decision-making problems subject to uncertainty, it may be essential to deal with multiple and often conflicting objectives while taking the decision-makers' risk preferences into account. Conditional value-at-risk…

Optimization and Control · Mathematics 2023-02-14 Najmesadat Nazemi , Sophie N. Parragh , Walter J. Gutjahr

Value-at-Risk (VaR) and Conditional Value-at-Risk (CVaR) are popular risk measures from academic, industrial and regulatory perspectives. The problem of minimizing CVaR is theoretically known to be of Neyman-Pearson type binary solution. We…

Portfolio Management · Quantitative Finance 2013-08-19 Jing Li , Mingxin Xu

Convex sample approximations of chance-constrained optimization problems are considered, in which chance constraints are replaced by sets of sampled constraints. We propose a randomized sample selection strategy that allows tight bounds to…

Optimization and Control · Mathematics 2018-05-22 Mark Cannon

Offline reinforcement learning (offline RL) algorithms often require additional constraints or penalty terms to address distribution shift issues, such as adding implicit or explicit policy constraints during policy optimization to reduce…

Machine Learning · Computer Science 2025-06-19 Ranting Hu

Scenario reduction (SR) alleviates the computational complexity of scenario-based stochastic optimization with conditional value-at-risk (SBSO-CVaR) by identifying representative scenarios to depict the underlying uncertainty and tail…

Optimization and Control · Mathematics 2025-10-20 Yingrui Zhuang , Lin Cheng , Ning Qi , Mads R. Almassalkhi , Feng Liu

This article develops a new algorithm named TTRISK to solve high-dimensional risk-averse optimization problems governed by differential equations (ODEs and/or PDEs) under uncertainty. As an example, we focus on the so-called Conditional…

Numerical Analysis · Mathematics 2022-12-02 Harbir Antil , Sergey Dolgov , Akwum Onwunta

In a wide variety of sequential decision making problems, it can be important to estimate the impact of rare events in order to minimize risk exposure. A popular risk measure is the conditional value-at-risk (CVaR), which is commonly…

Machine Learning · Statistics 2020-12-11 Dylan Troop , Frédéric Godin , Jia Yuan Yu

The geology of oil reservoirs is largely unknown. Consequently, the reservoir models used for production optimization are subject to significant uncertainty. To minimize the associated risk, the oil literature has mainly used ensemble-based…

Optimization and Control · Mathematics 2018-01-03 Andrea Capolei , Lasse Hjuler Christiansen , John Bagterp Jørgensen

We propose a sigmoidal approximation for the value-at-risk (that we call SigVaR) and we use this approximation to tackle nonlinear programs (NLPs) with chance constraints. We prove that the approximation is conservative and that the level…

Optimization and Control · Mathematics 2020-04-07 Yankai Cao , Victor M. Zavala

Many Machine Learning algorithms are formulated as regularized optimization problems, but their performance hinges on a regularization parameter that needs to be calibrated to each application at hand. In this paper, we propose a general…

Machine Learning · Statistics 2021-03-31 Mike Laszkiewicz , Asja Fischer , Johannes Lederer

Connected and automated vehicles (CAVs) provide the most intriguing opportunity to improve energy efficiency, traffic flow, and safety. In earlier work, we addressed the constrained optimal coordination problem of CAVs at different traffic…

Optimization and Control · Mathematics 2021-06-11 A M Ishtiaque Mahbub , Andreas A. Malikopoulos