中文
相关论文

相关论文: Acquisition of Project-Specific Assets with Bayesi…

200 篇论文

We investigate joint optimization on information acquisition and portfolio selection within a Bayesian adaptive framework. The investor dynamically controls the precision of a private signal and incurs costs while updating her belief about…

最优化与控制 · 数学 2025-08-19 Zongxia Liang , Shu Wang , Jianming Xia

We consider an infinite collection of agents who make decisions, sequentially, about an unknown underlying binary state of the world. Each agent, prior to making a decision, receives an independent private signal whose distribution depends…

计算机科学与博弈论 · 计算机科学 2012-09-07 Kimon Drakopoulos , Asuman Ozdaglar , John Tsitsiklis

We study a dynamic model of Bayesian persuasion in sequential decision-making settings. An informed principal observes an external parameter of the world and advises an uninformed agent about actions to take over time. The agent takes…

计算机科学与博弈论 · 计算机科学 2022-05-25 Jiarui Gan , Rupak Majumdar , Goran Radanovic , Adish Singla

In the Bayesian approach to sequential decision making, exact calculation of the (subjective) utility is intractable. This extends to most special cases of interest, such as reinforcement learning problems. While utility bounds are known to…

机器学习 · 计算机科学 2011-11-14 Christos Dimitrakakis

We model the joint distribution of choice probabilities and decision times in binary choice tasks as the solution to a problem of optimal sequential sampling, where the agent is uncertain of the utility of each action and pays a constant…

神经元与认知 · 定量生物学 2015-05-14 Drew Fudenberg , Philipp Strack , Tomasz Strzalecki

Research in reinforcement learning has produced algorithms for optimal decision making under uncertainty that fall within two main types. The first employs a Bayesian framework, where optimality improves with increased computational time.…

机器学习 · 统计学 2011-09-22 Christos Dimitrakakis

Any firm whose business strategy has an exposure constraint that limits its potential gain naturally considers expansion, as this can increase its exposure. We model business expansion as an enlargement of the opportunity set for business…

风险管理 · 定量金融 2021-12-14 Ling Wang , Kexin Chen , Mei Choi Chiu , Hoi Ying Wong

We establish explicit socially optimal rules for an irreversible investment deci- sion with time-to-build and uncertainty. Assuming a price sensitive demand function with a random intercept, we provide comparative statics and economic…

数理金融 · 定量金融 2014-06-03 René Aid , Salvatore Federico , Huyên Pham , Bertrand Villeneuve

In nonstationary bandit learning problems, the decision-maker must continually gather information and adapt their action selection as the latent state of the environment evolves. In each time period, some latent optimal action maximizes…

机器学习 · 计算机科学 2023-12-27 Seungki Min , Daniel Russo

This paper studies a sequential decision problem where payoff distributions are known and where the riskiness of payoffs matters. Equivalently, it studies sequential choice from a repeated set of independent lotteries. The decision-maker is…

理论经济学 · 经济学 2024-01-02 Zengjing Chen , Larry G. Epstein , Guodong Zhang

We study a problem of finding an optimal stopping strategy to liquidate an asset with unknown drift. Taking a Bayesian approach, we model the initial beliefs of an individual about the drift parameter by allowing an arbitrary probability…

数理金融 · 定量金融 2015-09-03 Erik Ekström , Juozas Vaicenavicius

Active learning is usually applied to acquire labels of informative data points in supervised learning, to maximize accuracy in a sample-efficient way. However, maximizing the accuracy is not the end goal when the results are used for…

The Markov Decision Process (MDP) is a popular framework for sequential decision-making problems, and uncertainty quantification is an essential component of it to learn optimal decision-making strategies. In particular, a Bayesian…

机器学习 · 统计学 2025-05-06 Jiaqi Guo , Chon Wai Ho , Sumeetpal S. Singh

Addressing uncertainty is critical for autonomous systems to robustly adapt to the real world. We formulate the problem of model uncertainty as a continuous Bayes-Adaptive Markov Decision Process (BAMDP), where an agent maintains a…

机器人学 · 计算机科学 2019-05-09 Gilwoo Lee , Brian Hou , Aditya Mandalika , Jeongseok Lee , Sanjiban Choudhury , Siddhartha S. Srinivasa

We consider an optimal stopping problem with n correlated offers where the goal is to design a (randomized) stopping strategy that maximizes the expected value of the offer in the sequence at which we stop. Instead of assuming to know the…

最优化与控制 · 数学 2025-07-08 Pieter Kleer , Daan Noordenbos

In this article, I investigate the use of Bayesian updating rules applied to modeling social agents in the case of continuos opinions models. Given another agent statement about the continuous value of a variable $x$, we will see that…

物理与社会 · 物理学 2009-04-04 Andre C. R. Martins

We develop a principled approach to end-to-end learning in stochastic optimization. First, we show that the standard end-to-end learning algorithm admits a Bayesian interpretation and trains a posterior Bayes action map. Building on the…

最优化与控制 · 数学 2023-06-13 Yves Rychener , Daniel Kuhn , Tobias Sutter

With the rise of the digital economy and an explosion of available information about consumers, effective personalization of goods and services has become a core business focus for companies to improve revenues and maintain a competitive…

机器学习 · 计算机科学 2022-11-04 Zhaonan Qu , Isabella Qian , Zhengyuan Zhou

We study the problem of learning to bid when the bidder's value is dynamic, i.e., when the current value depends on past outcomes. Specifically, we consider a bidder participating in repeated second-price auctions whose value depends on the…

机器学习 · 计算机科学 2026-05-28 Benjamin Heymann , Otmane Sakhi

We study the problem of choosing optimal policy rules in uncertain environments using models that may be incomplete and/or partially identified. We consider a policymaker who wishes to choose a policy to maximize a particular counterfactual…

计量经济学 · 经济学 2020-12-22 Thomas M. Russell