中文
相关论文

相关论文: A Bivariate Dead Band Process Adjustment Policy

200 篇论文

The Regression Discontinuity Design (RDD) is a quasi-experimental design that estimates the causal effect of a treatment when its assignment is defined by a threshold value for a continuous assignment variable. The RDD assumes that subjects…

应用统计 · 统计学 2020-03-27 Federico Ricciardi , Silvia Liverani , Gianluca Baio

This paper investigates the distributed optimal output consensus problem of second-order uncertain nonlinear multi-agent systems over weight-unbalanced directed networks. Under the standard assumption that local cost functions are strongly…

最优化与控制 · 数学 2022-02-11 Jin Zhang , Lu Liu , Haibo Ji

We consider the optimal control design problem for discrete-time LTI systems with state feedback, when the actuation signal is subject to unmeasurable switching propagation delays, due to e.g. the routing in a multi-hop communication…

系统与控制 · 计算机科学 2015-09-14 Antonio Cicone , Alessandro D'Innocenzo , Nicola Guglielmi , Linda Laglia

We consider the problem of discounted optimal state-feedback regulation for general unknown deterministic discrete-time systems. It is well known that open-loop instability of systems, non-quadratic cost functions and complex nonlinear…

系统与控制 · 电气工程与系统科学 2020-03-31 Alexandros Tanzanakis , John Lygeros

Battery participants in performance-based frequency regulation markets must consider the cost of battery aging in their operating strategies to maximize market profits. In this paper we solve this problem by proposing an optimal control…

最优化与控制 · 数学 2018-06-12 Bolun Xu , Yuanyuan Shi , Daniel S. Kirschen , Baosen Zhang

We consider a multi-armed bandit problem in a setting where each arm produces a noisy reward realization which depends on an observable random covariate. As opposed to the traditional static multi-armed bandit problem, this setting allows…

统计理论 · 数学 2013-05-27 Vianney Perchet , Philippe Rigollet

Load side participation can provide support to the power network by appropriately adapting the demand when required. In addition, it enables an economically improved power allocation. In this study, we consider the problem of providing an…

最优化与控制 · 数学 2021-05-10 Andreas Kasis , Stelios Timotheou , Marios Polycarpou

In many real-world reinforcement learning applications, access to the environment is limited to a fixed dataset, instead of direct (online) interaction with the environment. When using this data for either evaluation or training of a new…

机器学习 · 计算机科学 2019-11-06 Ofir Nachum , Yinlam Chow , Bo Dai , Lihong Li

We present a method for computing A-optimal sensor placements for infinite-dimensional Bayesian linear inverse problems governed by PDEs with irreducible model uncertainties. Here, irreducible uncertainties refers to uncertainties in the…

最优化与控制 · 数学 2020-08-26 Karina Koval , Alen Alexanderian , Georg Stadler

We analyze the control of the motion of a charged particle by means of an external electric field. The system is constrained to move along a given direction. The goal of the control is to change the speed of the particle in a fixed time…

最优化与控制 · 数学 2020-01-22 V. Martikyan , D. Guéry-Odelin , D. Sugny

In most machine learning training paradigms a fixed, often handcrafted, loss function is assumed to be a good proxy for an underlying evaluation metric. In this work we assess this assumption by meta-learning an adaptive loss function to…

Band selection refers to the process of choosing the most relevant bands in a hyperspectral image. By selecting a limited number of optimal bands, we aim at speeding up model training, improving accuracy, or both. It reduces redundancy…

图像与视频处理 · 电气工程与系统科学 2022-01-05 Lichao Mou , Sudipan Saha , Yuansheng Hua , Francesca Bovolo , Lorenzo Bruzzone , Xiao Xiang Zhu

This contribution considers one central aspect of experiment design in system identification. When a control design is based on an estimated model, the achievable performance is related to the quality of the estimate. The degradation in…

系统与控制 · 计算机科学 2013-03-22 Afrooz Ebadat , Mariette Annergren , Christian A. Larsson , Cristian R. Rojas , Bo Wahlberg

We study the problem of estimating the expected reward of the optimal policy in the stochastic disjoint linear bandit setting. We prove that for certain settings it is possible to obtain an accurate estimate of the optimal policy value even…

机器学习 · 计算机科学 2019-12-17 Weihao Kong , Gregory Valiant , Emma Brunskill

We present an algorithm for controlling and scheduling multiple linear time-invariant processes on a shared bandwidth limited communication network using adaptive sampling intervals. The controller is centralized and computes at every…

系统与控制 · 计算机科学 2015-06-25 Erik Henriksson , Daniel E. Quevedo , Edwin G. W. Peters , Henrik Sandberg , Karl Henrik Johansson

PID control architectures are widely used in industrial applications. Despite their low number of open parameters, tuning multiple, coupled PID controllers can become tedious in practice. In this paper, we extend PILCO, a model-based policy…

机器学习 · 计算机科学 2017-03-09 Andreas Doerr , Duy Nguyen-Tuong , Alonso Marco , Stefan Schaal , Sebastian Trimpe

The present paper deals with the problem of allocating patients to two competing treatments in the presence of covariates or prognostic factors in order to achieve a good trade-off among ethical concerns, inferential precision and…

统计方法学 · 统计学 2012-08-17 Alessandro Baldi Antognini , Maroussa Zagoraiou

We consider an optimal stochastic impulse control problem over an infinite time horizon motivated by a model of irreversible investment choices with fixed adjustment costs. By employing techniques of viscosity solutions and relying on…

最优化与控制 · 数学 2019-02-05 Salvatore Federico , Mauro Rosestolato , Elisa Tacconi

A new approach to solving two-point boundary value problems for a wave equation is developed. This new approach exploits the principle of stationary action to reformulate and solve such problems in the framework of optimal control. In…

最优化与控制 · 数学 2017-11-13 Peter M. Dower , William M. McEneaney

We consider the optimal design problem for a comparison of two regression curves, which is used to establish the similarity between the dose response relationships of two groups. An optimal pair of designs minimizes the width of the…

统计方法学 · 统计学 2014-11-19 Holger Dette , Kirsten Schorning