中文
相关论文

相关论文: Acquisition of Project-Specific Assets with Bayesi…

200 篇论文

The problem of making sequential decisions in unknown probabilistic environments is studied. In cycle $t$ action $y_t$ results in perception $x_t$ and reward $r_t$, where all quantities in general may depend on the complete history. The…

人工智能 · 计算机科学 2007-05-23 Marcus Hutter

We explore the question of how to learn an optimal search strategy within the example of a parking problem where parking opportunities arrive according to an unknown inhomogeneous Poisson process. The optimal policy is a threshold-type…

机器学习 · 计算机科学 2026-03-04 Stefan Ankirchner , Maximilian Philipp Thiel

In this paper we consider stopping problems with partial observation under a general risk-sensitive optimization criterion for problems with finite and infinite time horizon. Our aim is to maximize the certainty equivalent of the stopping…

最优化与控制 · 数学 2017-03-29 Nicole Bäuerle , Ulrich Rieder

We consider the batch (off-line) policy learning problem in the infinite horizon Markov Decision Process. Motivated by mobile health applications, we focus on learning a policy that maximizes the long-term average reward. We propose a…

统计理论 · 数学 2022-09-20 Peng Liao , Zhengling Qi , Runzhe Wan , Predrag Klasnja , Susan Murphy

In many applications, ranging from logistics to engineering, a designer is faced with a sequence of optimization tasks for which the objectives are in the form of black-box functions that are costly to evaluate. Furthermore, higher-fidelity…

机器学习 · 计算机科学 2025-01-09 Yunchuan Zhang , Sangwoo Park , Osvaldo Simeone

We study learning dynamics induced by strategic agents who repeatedly play a game with an unknown payoff-relevant parameter. In this dynamics, a belief estimate of the parameter is repeatedly updated given players' strategies and realized…

计算机科学与博弈论 · 计算机科学 2021-09-06 Manxi Wu , Saurabh Amin , Asuman Ozdaglar

We consider a real options model for the optimal irreversible investment problem of a profit maximizing company. The company has the opportunity to invest into a production plant capable of producing two products, of which the prices follow…

数理金融 · 定量金融 2021-07-09 Felix Dammann , Giorgio Ferrari

We study learning dynamics induced by strategic agents who repeatedly play a game with an unknown payoff-relevant parameter. In each step, an information system estimates a belief distribution of the parameter based on the players'…

系统与控制 · 电气工程与系统科学 2020-10-20 Manxi Wu , Saurabh Amin , Asuman Ozdaglar

Sequential Bayesian experimental design typically assumes that the number of experiments is fixed before data collection begins. In practical campaigns, however, experimentation may need to terminate early because additional measurements…

统计方法学 · 统计学 2026-05-29 Chen Cheng , Xun Huan

Portfolio construction is the science of balancing reward and risk; it is at the core of modern finance. In this paper, we tackle the question of optimal decision-making within a Bayesian paradigm, starting from a decision-theoretic…

应用统计 · 统计学 2024-11-12 Nicolas Nguyen , James Ridgway , Claire Vernade

Maximizing long-term rewards is the primary goal in sequential decision-making problems. The majority of existing methods assume that side information is freely available, enabling the learning agent to observe all features' states before…

机器学习 · 计算机科学 2023-07-19 Saeed Ghoorchian , Evgenii Kortukov , Setareh Maghsudi

In this article, a general problem of sequential statistical inference for general discrete-time stochastic processes is considered. The problem is to minimize an average sample number given that Bayesian risk due to incorrect decision does…

统计理论 · 数学 2010-10-18 Andrey Novikov

The model of a non-Bayesian agent who faces a repeated game with incomplete information against Nature is an appropriate tool for modeling general agent-environment interactions. In such a model the environment state (controlled by Nature)…

人工智能 · 计算机科学 2014-11-17 D. Monderer , M. Tennenholtz

A temporally abstract action, or an option, is specified by a policy and a termination condition: the policy guides option behavior, and the termination condition roughly determines its length. Generally, learning with longer options (like…

人工智能 · 计算机科学 2017-12-05 Anna Harutyunyan , Peter Vrancx , Pierre-Luc Bacon , Doina Precup , Ann Nowe

An autonomous experimentation platform in manufacturing is supposedly capable of conducting a sequential search for finding suitable manufacturing conditions by itself or even for discovering new materials with minimal human intervention.…

机器学习 · 计算机科学 2024-10-03 Imtiaz Ahmed , Satish Bukkapatnam , Bhaskar Botcha , Yu Ding

The timing of strategic exit is one of the most important but difficult business decisions, especially under competition and uncertainty. Motivated by this problem, we examine a stochastic game of exit in which players are uncertain about…

最优化与控制 · 数学 2023-10-09 H. Dharma Kwon , Jan Palczewski

This study proposes the General Bayes framework for policy learning. We consider decision problems in which a decision-maker chooses an action from an action set to maximize its expected welfare. Typical examples include treatment choice…

机器学习 · 统计学 2026-03-02 Masahiro Kato

A principal delegates choice to an agent whose decision depends on both beliefs and tastes. The principal can steer the delegated decision using two costly instruments: (i) an information policy that determines a Bayes--plausible…

理论经济学 · 经济学 2026-02-12 Kemal Ozbek

We consider the problem of sequential sampling from a finite number of independent statistical populations to maximize the expected infinite horizon average outcome per period, under a constraint that the expected average sampling cost does…

机器学习 · 统计学 2012-01-20 Apostolos Burnetas , Odysseas Kanavetas

When making decisions in a network, it is important to have up-to-date knowledge of the current state of the system. Obtaining this information, however, comes at a cost. In this paper, we determine the optimal finite-time update policy for…

信息论 · 计算机科学 2024-05-21 Eric Graves , Jake B. Perazzone , Kevin Chan