English
Related papers

Related papers: Finding the maximum-a-posteriori behaviour of agen…

200 papers

We consider the rates of noise-induced switching between the stable states of dissipative dynamical systems with delay and also the rates of noise-induced extinction, where such systems model population dynamics. We study a class of systems…

Statistical Mechanics · Physics 2015-01-27 Ira B. Schwartz , Lora Billings , Thomas W. Carr , Mark Dykman

We theoretically and numerically study the problem of optimal control of large-scale autonomous systems under explicitly adversarial conditions, including probabilistic destruction of agents during the simulation. Large-scale autonomous…

Optimization and Control · Mathematics 2021-08-06 Theodoros Tsatsanifos , Abram H. Clark , Claire Walton , Isaac Kaminer , Qi Gong

Online decision-making can be formulated as the popular stochastic multi-armed bandit problem where a learner makes decisions (or takes actions) to maximize cumulative rewards collected from an unknown environment. This paper proposes to…

Systems and Control · Electrical Eng. & Systems 2025-11-26 Jonathan Gornet , Mehdi Hosseinzadeh , Bruno Sinopoli

AI planning algorithms have addressed the problem of generating sequences of operators that achieve some input goal, usually assuming that the planning agent has perfect control over and information about the world. Relaxing these…

Artificial Intelligence · Computer Science 2013-02-28 Denise L. Draper , Steve Hanks , Daniel Weld

In dynamic programming and reinforcement learning, the policy for the sequential decision making of an agent in a stochastic environment is usually determined by expressing the goal as a scalar reward function and seeking a policy that…

Artificial Intelligence · Computer Science 2025-02-26 Simon Dima , Simon Fischer , Jobst Heitzig , Joss Oliver

Time-varying systems are a challenge in many scientific and engineering areas. Usually, estimation of time-varying parameters or signals must be performed online, which calls for the development of responsive online algorithms. In this…

Optimization and Control · Mathematics 2018-09-10 Sophie M. Fosson

In this paper, we disclose the statistical behavior of the max-product algorithm configured to solve a maximum a posteriori (MAP) estimation problem in a network of distributed agents. Specifically, we first build a distributed hypothesis…

Information Theory · Computer Science 2020-04-30 Younes Abdi , Tapani Ristaniemi

Motivated by the need to secure cyber-physical systems against attacks, we consider the problem of estimating the state of a noisy linear dynamical system when a subset of sensors is arbitrarily corrupted by an adversary. We propose a…

Optimization and Control · Mathematics 2015-04-24 Shaunak Mishra , Yasser Shoukry , Nikhil Karamchandani , Suhas Diggavi , Paulo Tabuada

Multi-agent reinforcement learning systems aim to provide interacting agents with the ability to collaboratively learn and adapt to the behaviour of other agents. In many real-world applications, the agents can only acquire a partial view…

Machine Learning · Computer Science 2018-12-04 Ozsel Kilinc , Giovanni Montana

Data-driven, model-free analytics are natural choices for discovery and forecasting of complex, nonlinear systems. Methods that operate in the system state-space require either an explicit multidimensional state-space, or, one approximated…

Machine Learning · Statistics 2021-03-15 Joseph Park , Gerald M Pao , Erik Stabenau , George Sugihara , Thomas Lorimer

This paper analyzes a service system modeled as a single-server queue, in which the service provider aims to dynamically maximize the expected revenue per unit of time. This is achieved by constructing a stochastic gradient descent…

Optimization and Control · Mathematics 2026-03-05 Shreehari Anand Bodas , Harsha Honnappa , Michel Mandjes , Liron Ravner

In the context of high-dimensional linear regression models, we propose an algorithm of exact support recovery in the setting of noisy compressed sensing where all entries of the design matrix are independent and identically distributed…

Statistics Theory · Mathematics 2019-10-23 Mohamed Ndaoud , Alexandre B. Tsybakov

This paper deals with the problem of formulating an adaptive Model Predictive Control strategy for constrained uncertain systems. We consider a linear system, in presence of bounded time varying additive uncertainty. The uncertainty is…

Systems and Control · Electrical Eng. & Systems 2021-04-13 Monimoy Bujarbaruah , Xiaojing Zhang , Marko Tanaskovic , Francesco Borrelli

We derive a new method to infer from data the out-of-equilibrium alignment dynamics of collectively moving animal groups, by considering the maximum entropy distribution consistent with temporal and spatial correlations of flight direction.…

Multi-agent learning is a challenging problem in machine learning that has applications in different domains such as distributed control, robotics, and economics. We develop a prescriptive model of multi-agent behavior using Markov games.…

Artificial Intelligence · Computer Science 2020-05-27 Jalal Etesami , Christoph-Nikolas Straehle

We study decision timing problems on finite horizon with Poissonian information arrivals. In our model, a decision maker wishes to optimally time her action in order to maximize her expected reward. The reward depends on an unobservable…

Optimization and Control · Mathematics 2012-05-07 Michael Ludkovski , Semih Sezer

We present a full stochastic description of the pair approximation scheme to study binary-state dynamics on heterogeneous networks. Within this general approach, we obtain a set of equations for the dynamical correlations, fluctuations and…

Physics and Society · Physics 2018-11-05 A. F. Peralta , A. Carro , M. San Miguel , R. Toral

A new class of stochastic processes called independent and periodically identically distributed (i.p.i.d.) processes is defined to capture periodically varying statistical behavior. Algorithms are proposed to detect changes in such i.p.i.d.…

Statistics Theory · Mathematics 2018-10-31 Taposh Banerjee , Prudhvi Gurram , Gene Whipps

We develop a formalism to describe the discrete-time dynamics of systems containing an arbitrary number of interacting species. The individual-based model, which forms our starting point, is described by a Markov chain, which in the limit…

Statistical Mechanics · Physics 2014-10-06 César Parra-Rojas , Joseph D. Challenger , Duccio Fanelli , Alan J. McKane

We address the problem of parameter estimation in models of systems biology from noisy observations. The models we consider are characterized by simultaneous deterministic nonlinear differential equations whose parameters are either taken…

Machine Learning · Statistics 2017-05-01 Xin Liu , Mahesan Niranjan