中文
相关论文

相关论文: Dynamic Non-Bayesian Decision Making

200 篇论文

In nonstationary bandit learning problems, the decision-maker must continually gather information and adapt their action selection as the latent state of the environment evolves. In each time period, some latent optimal action maximizes…

机器学习 · 计算机科学 2023-12-27 Seungki Min , Daniel Russo

This paper studies a nonlinear filtering problem over an infinite time interval. The signal to be estimated is driven by a stochastic partial differential equation involves unknown parameters. Based on discrete observation, strongly…

统计理论 · 数学 2021-07-12 Qizhu Liang , Jie Xiong , Xingqiu Zhao

Consider a team of agents in the plane searching for and visiting target points that appear in a bounded environment according to a stochastic renewal process with a known absolutely continuous spatial distribution. Agents must detect…

机器人学 · 计算机科学 2012-02-02 John J. Enright , Emilio Frazzoli

We consider a simulation optimization problem for a context-dependent decision-making, which aims to determine the top-m designs for all contexts. Under a Bayesian framework, we formulate the optimal dynamic sampling decision as a…

机器学习 · 统计学 2023-06-12 Gongbo Zhang , Sihua Chen , Kuihua Huang , Yijie Peng

Humans and animals explore their environment and acquire useful skills even in the absence of clear goals, exhibiting intrinsic motivation. The study of intrinsic motivation in artificial agents is concerned with the following question:…

Large dynamic economies with heterogeneous agents and aggregate shocks are central to many important applications, yet their equilibrium analysis remains computationally challenging. This is because the standard solution approach, rational…

综合经济学 · 经济学 2025-02-25 Bilal Islah , Bar Light

Is there a canonical way to think of agency beyond reward maximisation? In this paper, we show that any type of behaviour complying with physically sound assumptions about how macroscopic biological agents interact with the world…

人工智能 · 计算机科学 2024-01-24 Lancelot Da Costa , Samuel Tenka , Dominic Zhao , Noor Sajid

In this paper, we describe an integrated framework for autonomous decision making in a dynamic and interactive environment. We model the interactions between the ego agent and its operating environment as a two-player dynamic game, and…

人工智能 · 计算机科学 2019-09-19 Sisi Li , Nan Li , Anouck Girard , Ilya Kolmanovsky

We investigate the problem of persistently monitoring a finite set of targets with internal states that evolve with linear stochastic dynamics using a finite set of mobile agents. We approach the problem from the infinite-horizon…

系统与控制 · 电气工程与系统科学 2019-10-01 Samuel C. Pinto , Sean B. Andersson , Julien M. Hendrickx , Christos G. Cassandras

Engineers are often faced with the decision to select the most appropriate model for simulating the behavior of engineered systems, among a candidate set of models. Experimental monitoring data can generate significant value by supporting…

应用统计 · 统计学 2023-10-17 Antonios Kamariotis , Eleni Chatzi

A default assumption in the design of reinforcement-learning algorithms is that a decision-making agent always explores to learn optimal behavior. In sufficiently complex environments that approach the vastness and scale of the real world,…

机器学习 · 计算机科学 2024-07-23 Dilip Arumugam , Saurabh Kumar , Ramki Gummadi , Benjamin Van Roy

Multi-Agent Reinforcement Learning involves agents that learn together in a shared environment, leading to emergent dynamics sensitive to initial conditions and parameter variations. A Dynamical Systems approach, which studies the evolution…

多智能体系统 · 计算机科学 2025-01-03 David Goll , Jobst Heitzig , Wolfram Barfuss

Decision makers often aim to learn a treatment assignment policy under a capacity constraint on the number of agents that they can treat. When agents can respond strategically to such policies, competition arises, complicating estimation of…

机器学习 · 统计学 2025-03-31 Roshni Sahoo , Stefan Wager

We develop a learning-based algorithm for the distributed formation control of networked multi-agent systems governed by unknown, nonlinear dynamics. Most existing algorithms either assume certain parametric forms for the unknown dynamic…

系统与控制 · 电气工程与系统科学 2022-01-13 Christos K. Verginis , Zhe Xu , Ufuk Topcu

The socioeconomic impact of pollution naturally comes with uncertainty due to, e.g., current new technological developments in emissions' abatement or demographic changes. On top of that, the trend of the future costs of the environmental…

最优化与控制 · 数学 2024-02-28 Matteo Basei , Giorgio Ferrari , Neofytos Rodosthenous

Continuous control of non-stationary environments is a major challenge for deep reinforcement learning algorithms. The time-dependency of the state transition dynamics aggravates the notorious stability problems of model-free deep…

机器学习 · 计算机科学 2025-11-05 Abdullah Akgül , Gulcin Baykal , Manuel Haußmann , Melih Kandemir

We study the problem of matching agents who arrive at a marketplace over time and leave after d time periods. Agents can only be matched while they are present in the marketplace. Each pair of agents can yield a different match value, and…

数据结构与算法 · 计算机科学 2018-03-06 Itai Ashlagi , Maximilien Burq , Patrick Jaillet , Amin Saberi

When robots share the same workspace with other intelligent agents (e.g., other robots or humans), they must be able to reason about the behaviors of their neighboring agents while accomplishing the designated tasks. In practice,…

机器人学 · 计算机科学 2022-10-18 Junhong Xu , Durgakant Pushp , Kai Yin , Lantao Liu

We use model-free reinforcement learning, extensive simulation, and transfer learning to develop a continuous control algorithm that has good zero-shot performance in a real physical environment. We train a simulated agent to act optimally…

人工智能 · 计算机科学 2018-03-09 M Ferguson , K. H. Law

Two traditional paradigms are often used to describe the behavior of agents in multi-agent complex systems. In the first one, agents are considered to be fully rational and systems are seen as multi-player games. In the second one, agents…

计算机科学与博弈论 · 计算机科学 2016-03-17 Mickael Randour