English
Related papers

Related papers: Gradient-based Cooperative Control of quasi-Linear…

200 papers

We address the traffic light control problem for a single intersection by viewing it as a stochastic hybrid system and developing a Stochastic Flow Model (SFM) for it. We adopt a quasi-dynamic control policy based on partial state…

Systems and Control · Computer Science 2013-08-06 Yanfeng Geng , Christos G. Cassandras

A high-gain observer-based cooperative deterministic learning (CDL) control algorithm is proposed in this chapter for a group of identical unicycle-type unmanned ground vehicles (UGVs) to track over desired reference trajectories. For the…

Systems and Control · Electrical Eng. & Systems 2020-02-25 Xiaonan Dong , Paolo Stegagno , Chengzhi Yuan , Wei Zeng

A wide range of multi-agent coordination problems including reference tracking and disturbance rejection requirements can be formulated as a cooperative output regulation problem. The general framework captures typical problems such as…

Systems and Control · Computer Science 2014-06-04 Georg Seyboth , Wei Ren , Frank Allgöwer

Gradient matching is a promising tool for learning parameters and state dynamics of ordinary differential equations. It is a grid free inference approach, which, for fully observable systems is at times competitive with numerical…

Machine Learning · Statistics 2018-04-11 Nico S. Gorbach , Stefan Bauer , Joachim M. Buhmann

Instrumental variables (IVs) provide a powerful strategy for identifying causal effects in the presence of unobservable confounders. Within the nonparametric setting (NPIV), recent methods have been based on nonlinear generalizations of…

Machine Learning · Statistics 2024-12-24 Yuri Fonseca , Caio Peixoto , Yuri Saporito

We propose a method for finding approximate compilations of quantum unitary transformations, based on techniques from policy gradient reinforcement learning. The choice of a stochastic policy allows us to rephrase the optimization problem…

Quantum Physics · Physics 2022-09-14 David A. Herrera-Martí

Cooperative Adaptive Cruise Control (CACC) enables vehicle platooning through inter-vehicle communication, improving traffic efficiency and safety. Conventional CACC relies on feedback linearization, assuming exact vehicle parameters;…

Systems and Control · Electrical Eng. & Systems 2026-02-12 Mischa Huisman , Thomas Arnold , Erjen Lefeber , Nathan van de Wouw , Carlos Murguia

Policy gradient is a generic and flexible reinforcement learning approach that generally enjoys simplicity in analysis, implementation, and deployment. In the last few decades, this approach has been extensively advanced for fully…

Machine Learning · Computer Science 2020-05-26 Kamyar Azizzadenesheli , Yisong Yue , Animashree Anandkumar

The task of estimating the gradient of a function in the presence of noise is central to several forms of reinforcement learning, including policy search methods. We present two techniques for reducing gradient estimation errors in the…

Machine Learning · Computer Science 2012-12-12 Gregory Lawrence , Noah Cowan , Stuart Russell

It seems that in the current age, computers, computation, and data have an increasingly important role to play in scientific research and discovery. This is reflected in part by the rise of machine learning and artificial intelligence,…

Machine Learning · Computer Science 2024-05-15 Ronan Keane

We propose the use of latent space generative world models to address the covariate shift problem in autonomous driving. A world model is a neural network capable of predicting an agent's next state given past states and actions. By…

In this paper, we study the control of dynamical systems under temporal logic task specifications using gradient-based methods relying on quantitative measures that express the extent to which the tasks are satisfied. A class of controllers…

Systems and Control · Electrical Eng. & Systems 2019-09-06 Peter Varnai , Dimos V. Dimarogonas

In this paper, we study the gradient descent-ascent method for convex-concave saddle-point problems. We derive a new non-asymptotic global convergence rate in terms of distance to the solution set by using the semidefinite programming…

Optimization and Control · Mathematics 2022-09-19 Moslem Zamani , Hadi Abbaszadehpeivasti , Etienne de Klerk

We consider the problem of direct data-driven predictive control for unknown stochastic linear time-invariant (LTI) systems with partial state observation. Building upon our previous research on data-driven stochastic control, this paper…

Systems and Control · Electrical Eng. & Systems 2024-09-12 Ruiqi Li , John W. Simpson-Porco , Stephen L. Smith

This paper proposes a cooperative strategy of connected and automated vehicles (CAVs) longitudinal control for partially connected and automated traffic environment based on deep reinforcement learning (DRL) algorithm, which enhances the…

Systems and Control · Electrical Eng. & Systems 2020-12-04 Haotian Shi , Yang Zhou , Keshu Wu , Xin Wang , Yangxin Lin , Bin Ran

This work establishes a general stochastic maximum principle for partially observed optimal control of semi-linear stochastic partial differential equations in a nonconvex control domain. The state evolves in a Hilbert space driven by a…

Optimization and Control · Mathematics 2025-04-22 Yanzhao Cao , Hongjiang Qian , George Yin

Variational inference in Bayesian deep learning often involves computing the gradient of an expectation that lacks a closed-form solution. In these cases, pathwise and score-function gradient estimators are the most common approaches. The…

Machine Learning · Statistics 2024-10-10 Kenyon Ng , Susan Wei

In this paper, we investigate a model-free optimal control design that minimizes an infinite horizon average expected quadratic cost of states and control actions subject to a probabilistic risk or chance constraint using input-output data.…

Systems and Control · Electrical Eng. & Systems 2024-11-11 Arunava Naha , Subhrakanti Dey

Many recently proposed gradient projection algorithms with inertial extrapolation step for solving quasi-variational inequalities in Hilbert spaces are proven to be strongly convergent with no linear rate given when the cost operator is…

Optimization and Control · Mathematics 2024-04-23 Yonghong Yao , Lateef O. Jolaoso , Yekini Shehu

Optimal tracking of continuous time nonlinear systems has been extensively studied in literature. However, in several applications, absence of knowledge about system dynamics poses a severe challenge to solving the optimal tracking problem.…

Systems and Control · Electrical Eng. & Systems 2020-01-22 Amardeep Mishra , Satadal Ghosh
‹ Prev 1 4 5 6 7 8 10 Next ›