中文
相关论文

相关论文: Dynamic Non-Bayesian Decision Making

200 篇论文

I study a principal-agent model in which a principal hires an agent to collect information about an unknown continuous state. The agent acquires a signal whose distribution is centered around the state, controlling the signal's precision at…

理论经济学 · 经济学 2026-05-05 Fan Wu

We consider a sequence of repeated interactions between an agent and an environment. Uncertainty about the environment is captured by a probability distribution over a space of hypotheses, which includes all computable functions. Given a…

人工智能 · 计算机科学 2009-12-02 Peter de Blanc

We address the problem where a mobile search agent seeks to find an unknown number of stationary objects distributed in a bounded search domain, and the search mission is subject to time/distance constraint. Our work accounts for false…

机器人学 · 计算机科学 2018-06-26 Harun Yetkin , Collin Lutz , Daniel Stilwell

The aim of a number of psychophysics tasks is to uncover how mammals make decisions in a world that is in flux. Here we examine the characteristics of ideal and near-ideal observers in a task of this type. We ask when and how performance…

神经元与认知 · 定量生物学 2019-10-10 Adrian E. Radillo , Alan Veliz-Cuba , Krešimir Josić , Zachary P. Kilpatrick

This paper investigates the dynamics of noncooperative interactions between artificial intelligence agents and human decision-makers in strategic environments. In particular, motivated by extensive literature in behavioral Economics, human…

计算机科学与博弈论 · 计算机科学 2026-03-19 Dylan Waldner , Vyacheslav Kungurtsev , Mitchelle Ashimosi

Experiments in engineering are typically conducted in controlled environments where parameters can be set to any desired value. This assumes that the same applies in a real-world setting -- an assumption that is often incorrect as many…

机器学习 · 计算机科学 2025-11-18 Mike Diessner , Kevin J. Wilson , Richard D. Whalley

We consider the problem of designing a sequential decision making agent to maximize an unknown time-varying function which switches with time. At each step, the agent receives an observation of the function's value at a point decided by the…

最优化与控制 · 数学 2023-11-07 Durgesh Kalwar , Vineeth B. S

Policy evaluation estimates the performance of a policy by (1) collecting data from the environment and (2) processing raw data into a meaningful estimate. Due to the sequential nature of reinforcement learning, any improper data-collecting…

机器学习 · 计算机科学 2025-03-21 Shuze Daniel Liu , Claire Chen , Shangtong Zhang

We develop a hierarchical Bayesian dynamic game for competitive inventory and pricing under incomplete information. Two firms repeatedly choose order quantities and prices while facing two layers of uncertainty: unknown market demand and…

统计方法学 · 统计学 2026-03-09 Debashis Chatterjee

Bounded rational agents often make decisions by evaluating a finite selection of choices, typically derived from a reference point termed the $`$default policy,' based on previous experience. However, the inherent rigidity of the static…

机器人学 · 计算机科学 2024-09-19 Durgakant Pushp , Junhong Xu , Zheng Chen , Lantao Liu

Present bias, the tendency to overvalue immediate rewards while undervaluing future ones, is a well-known barrier to achieving long-term goals. As artificial intelligence and behavioral economics increasingly focus on this phenomenon, the…

计算机科学与博弈论 · 计算机科学 2024-09-18 Yasunori Akagi , Hideaki Kim , Takeshi Kurashima

Subject to reasonable conditions, in large population stochastic dynamics games, where the agents are coupled by the system's mean field (i.e. the state distribution of the generic agent) through their nonlinear dynamics and their nonlinear…

最优化与控制 · 数学 2019-05-28 Nevroz Sen , Peter E. Caines

We propose a model of inference and heuristic decision-making in groups that is rooted in the Bayes rule but avoids the complexities of rational inference in partially observed environments with incomplete information, which are…

多智能体系统 · 计算机科学 2016-11-04 M. Amin Rahimian , Ali Jadbabaie

We consider a ubiquitous scenario in the Internet economy when individual decision-makers (henceforth, agents) both produce and consume information as they make strategic choices in an uncertain environment. This creates a three-way…

计算机科学与博弈论 · 计算机科学 2021-04-09 Yishay Mansour , Aleksandrs Slivkins , Vasilis Syrgkanis , Zhiwei Steven Wu

Selective classification is a powerful tool for automated decision-making in high-risk scenarios, allowing classifiers to act only when confident and abstain when uncertainty is high. Given a target accuracy, our goal is to minimize…

统计理论 · 数学 2025-10-28 Mohamed Ndaoud , Peter Radchenko , Bradley Rava

An adaptive agent predicting the future state of an environment must weigh trust in new observations against prior experiences. In this light, we propose a view of the adaptive immune system as a dynamic Bayesian machinery that updates its…

种群与进化 · 定量生物学 2019-05-14 Andreas Mayer , Vijay Balasubramanian , Aleksandra M. Walczak , Thierry Mora

Model-based reinforcement learning has the potential to be more sample efficient than model-free approaches. However, existing model-based methods are vulnerable to model bias, which leads to poor generalization and asymptotic performance…

机器学习 · 计算机科学 2019-06-27 Tung-Long Vuong , Kenneth Tran

We consider the problem of learning by demonstration from agents acting in unknown stochastic Markov environments or games. Our aim is to estimate agent preferences in order to construct improved policies for the same task that the agents…

机器学习 · 计算机科学 2014-08-12 Aristide Tossou , Christos Dimitrakakis

We consider the problem of learning by demonstration from agents acting in unknown stochastic Markov environments or games. Our aim is to estimate agent preferences in order to construct improved policies for the same task that the agents…

机器学习 · 统计学 2013-07-16 Aristide C. Y. Tossou , Christos Dimitrakakis

We investigate the matching of agents to resources in a computational ecology configured to present heterogeneous resource patches to evolving, neurally controlled agents. We repeatedly find a nearly optimal, ideal free distribution (IFD)…

种群与进化 · 定量生物学 2011-12-16 Virgil Griffith , Larry S. Yaeger
‹ 上一页 1 8 9 10 下一页 ›