English
Related papers

Related papers: Episodic Bayesian Optimal Control with Unknown Ran…

200 papers

This PhD thesis presents a distributional view of optimization in place of a worst-case perspective. We motivate this view with an investigation of the failure point of classical optimization. Subsequently we consider the optimization of a…

Optimization and Control · Mathematics 2025-07-23 Felix Benning

Addressing uncertainty is critical for autonomous systems to robustly adapt to the real world. We formulate the problem of model uncertainty as a continuous Bayes-Adaptive Markov Decision Process (BAMDP), where an agent maintains a…

This paper develops a unified methodology for probabilistic analysis and optimal control design for jump diffusion processes defined by polynomials. For such systems, the evolution of the moments of the state can be described via a system…

Optimization and Control · Mathematics 2017-02-03 Andrew Lamperski , Khem Raj Ghusinga , Abhyudai Singh

We consider a stochastic control problem for a class of nonlinear kernels. More precisely, our problem of interest consists in the optimisation, over a set of possibly non-dominated probability measures, of solutions of backward stochastic…

Probability · Mathematics 2017-07-28 Dylan Possamaï , Xiaolu Tan , Chao Zhou

In this paper, we consider the classic stochastic (dynamic) knapsack problem, a fundamental mathematical model in revenue management, with general time-varying random demand. Our main goal is to study the optimal policies, which can be…

Optimization and Control · Mathematics 2018-07-19 Yingdong Lu

A new stochastic control problem of population dynamics under partial observation is formulated and analyzed both mathematically and numerically, with an emphasis on environmental and ecological problems. The decision-maker can only…

Optimization and Control · Mathematics 2020-04-13 Hidekazu Yoshioka , Yuta Yaegashi , Motoh Tsujimura

The solution to a stochastic optimal control problem can be determined by computing the value function from a discretization of the associated Hamilton-Jacobi-Bellman equation. Alternatively, the problem can be reformulated in terms of a…

Optimization and Control · Mathematics 2024-02-29 Sebastian Reich

The problem of state estimation for unobservable distribution systems is considered. A deep learning approach to Bayesian state estimation is proposed for real-time applications. The proposed technique consists of distribution learning of…

Machine Learning · Statistics 2019-02-26 Kursat Rasim Mestav , Jaime Luengo-Rozas , Lang Tong

Without exact knowledge of the true system dynamics, optimal control of non-linear continuous-time systems requires careful treatment under epistemic uncertainty. In this work, we translate a probabilistic interpretation of the Pontryagin…

Machine Learning · Computer Science 2025-09-03 David Leeftink , Çağatay Yıldız , Steffen Ridderbusch , Max Hinne , Marcel van Gerven

We study high-dimensional stochastic optimal control problems in which many agents cooperate to minimize a convex cost functional. We consider both the full-information problem, in which each agent observes the states of all other agents,…

Probability · Mathematics 2023-01-10 Joe Jackson , Daniel Lacker

Bayesian optimization has become a popular method for high-throughput computing, like the design of computer experiments or hyperparameter tuning of expensive models, where sample efficiency is mandatory. In these applications, distributed…

Machine Learning · Computer Science 2019-07-08 Javier Garcia-Barcos , Ruben Martinez-Cantin

This paper deals with the problem of asymptotically optimal detection of changes in regime-switching stochastic models. We need to divide the whole obtained sample of data into several sub-samples with observations belonging to different…

Statistics Theory · Mathematics 2013-01-25 Boris Brodsky , Boris Darkhovsky

Non-parametric episodic memory can be used to quickly latch onto high-rewarded experience in reinforcement learning tasks. In contrast to parametric deep reinforcement learning approaches in which reward signals need to be back-propagated…

Machine Learning · Computer Science 2023-04-25 Zhao Yang , Thomas M. Moerland , Mike Preuss , Aske Plaat

A new class of stochastic processes called independent and periodically identically distributed (i.p.i.d.) processes is defined to capture periodically varying statistical behavior. Algorithms are proposed to detect changes in such i.p.i.d.…

Statistics Theory · Mathematics 2018-10-31 Taposh Banerjee , Prudhvi Gurram , Gene Whipps

Sequential Model-based Bayesian Optimization has been successful-ly applied to several application domains, characterized by complex search spaces, such as Automated Machine Learning and Neural Architecture Search. This paper focuses on…

Systems and Control · Electrical Eng. & Systems 2020-03-10 Antonio Candelieri , Bruno Galuzzi , Ilaria Giordani , Francesco Archetti

In this paper, we study a data-enabled predictive control (DeePC) algorithm applied to unknown stochastic linear time-invariant systems. The algorithm uses noise-corrupted input/output data to predict future trajectories and compute optimal…

Optimization and Control · Mathematics 2019-11-04 Jeremy Coulson , John Lygeros , Florian Dörfler

This article presents a constrained policy optimization approach for the optimal control of systems under nonstationary uncertainties. We introduce an assumption that we call Markov embeddability that allows us to cast the stochastic…

Optimization and Control · Mathematics 2026-05-11 Sungho Shin , François Pacaud , Emil Contantinescu , Mihai Anitescu

The paper solves the problem of optimal portfolio choice when the parameters of the asset returns distribution, like the mean vector and the covariance matrix are unknown and have to be estimated by using historical data of the asset…

Statistical Finance · Quantitative Finance 2023-04-19 David Bauder , Taras Bodnar , Nestor Parolya , Wolfgang Schmid

In this paper, we consider a class of stochastic control problems for stochastic differential equations with random coefficients. The control domain need not to be convex but the control process is not allowed to enter in diffusion term.…

Optimization and Control · Mathematics 2020-08-06 Ishak Alia , Mohamed Sofiane Alia

This paper presents a general asymptotic theory of sequential Bayesian estimation giving results for the strongest, almost sure convergence. We show that under certain smoothness conditions on the probability model, the greedy information…

Statistics Theory · Mathematics 2016-01-11 Janne V. Kujala
‹ Prev 1 4 5 6 7 8 10 Next ›