中文
相关论文

相关论文: Acquisition of Project-Specific Assets with Bayesi…

200 篇论文

We consider sequential decision problems in which we adaptively choose one of finitely many alternatives and observe a stochastic reward. We offer a new perspective of interpreting Bayesian ranking and selection problems as adaptive…

机器学习 · 计算机科学 2016-06-16 Yingfei Wang , Warren Powell

Bayesian decision theory outlines a rigorous framework for making optimal decisions based on maximizing expected utility over a model posterior. However, practitioners often do not have access to the full posterior and resort to approximate…

机器学习 · 统计学 2019-10-29 Tomasz Kuśmierczyk , Joseph Sakaya , Arto Klami

Even in the face of deteriorating and highly volatile demand, firms often invest in, rather than discard, aging technologies. In order to study this phenomenon, we model the firm's profit stream as a Brownian motion with negative drift. At…

最优化与控制 · 数学 2019-01-08 H. Dharma Kwon

A decision-maker periodically acquires information about a changing state, controlling both the timing and content of updates. I characterize optimal policies using a decomposition of the dynamic problem into optimal stopping and static…

理论经济学 · 经济学 2025-12-02 César Barilla

The socioeconomic impact of pollution naturally comes with uncertainty due to, e.g., current new technological developments in emissions' abatement or demographic changes. On top of that, the trend of the future costs of the environmental…

最优化与控制 · 数学 2024-02-28 Matteo Basei , Giorgio Ferrari , Neofytos Rodosthenous

We define an online learning and optimization problem with discrete and irreversible decisions contributing toward a coverage target. In each period, a decision-maker selects facilities to open, receives information on the success of each…

机器学习 · 计算机科学 2026-03-06 Alexandre Jacquillat , Michael Lingzhi Li

We study decision timing problems on finite horizon with Poissonian information arrivals. In our model, a decision maker wishes to optimally time her action in order to maximize her expected reward. The reward depends on an unobservable…

最优化与控制 · 数学 2012-05-07 Michael Ludkovski , Semih Sezer

An agent acquires information dynamically until her belief about a binary state reaches an upper or lower threshold. She can choose any signal process subject to a constraint on the rate of entropy reduction. Strategies are ordered by "time…

理论经济学 · 经济学 2024-08-23 Daniel Chen , Weijie Zhong

We develop a Bayesian model for decision-making under time pressure with endogenous information acquisition. In our model, the decision maker decides when to observe (costly) information by sampling an underlying continuous-time stochastic…

人工智能 · 计算机科学 2016-10-25 Ahmed M. Alaa , Mihaela van der Schaar

We consider a model where an agent has a repeated decision to make and wishes to maximize their total payoff. Payoffs are influenced by an action taken by the agent, but also an unknown state of the world that evolves over time. Before…

计算机科学与博弈论 · 计算机科学 2021-01-20 Nicole Immorlica , Ian Kash , Brendan Lucier

We consider the problem of sequentially making decisions that are rewarded by "successes" and "failures" which can be predicted through an unknown relationship that depends on a partially controllable vector of attributes for each instance.…

机器学习 · 统计学 2017-09-18 Yingfei Wang , Chu Wang , Warren Powell

We study a model of irreversible investment for a decision-maker who has the possibility to gradually invest in a project with unknown value. In this setting, we introduce and explore a feature of "learning-by-doing", where the learning…

最优化与控制 · 数学 2024-06-25 Erik Ekström , Yerkin Kitapbayev , Alessandro Milazzo , Topias Tolonen-Weckström

This paper investigates the investment problem of constructing an optimal no-short sequential portfolio strategy in a market with a latent dependence structure between asset prices and partly unobservable side information, which is often…

数理金融 · 定量金融 2025-01-22 Duy Khanh Lam

We develop a hierarchical Bayesian dynamic game for competitive inventory and pricing under incomplete information. Two firms repeatedly choose order quantities and prices while facing two layers of uncertainty: unknown market demand and…

统计方法学 · 统计学 2026-03-09 Debashis Chatterjee

We exhibit optimal control strategies for a simple toy problem in which the underlying dynamics depend on a parameter that is initially unknown and must be learned. We consider a cost function posed over a finite time interval, in contrast…

最优化与控制 · 数学 2020-02-27 Charles L. Fefferman , Bernat Guillen Pegueroles , Clarence W. Rowley , Melanie Weber

This paper studies optimal consumption and saving decisions under uncertainty about the transition dynamics of the economic environment. We consider a general optimal savings problem in which the exogenous state governing discounting,…

理论经济学 · 经济学 2026-03-10 Qingyin Ma , Xinxin Zhang

The exploration-exploitation trade-off is among the central challenges of reinforcement learning. The optimal Bayesian solution is intractable in general. This paper studies to what extent analytic statements about optimal learning are…

机器学习 · 统计学 2015-03-13 Philipp Hennig

Linear dynamical systems that obey stochastic differential equations are canonical models. While optimal control of known systems has a rich literature, the problem is technically hard under model uncertainty and there are hardly any…

系统与控制 · 电气工程与系统科学 2023-06-09 Mohamad Kazem Shirani Faradonbeh , Mohamad Sadegh Shirani Faradonbeh

We analyze an irreversible investment decision for a project which yields a flow of future operating profits given by a geometric Brownian motion with unknown drift. In contrast to similar optimal stopping problems with incomplete…

最优化与控制 · 数学 2025-02-19 Fabian Gierens , Berenice Anne Neumann

Real-time inference is a challenge of real-world reinforcement learning due to temporal differences in time-varying environments: the system collects data from the past, updates the decision model in the present, and deploys it in the…

机器学习 · 计算机科学 2024-05-28 Hyunin Lee , Ming Jin , Javad Lavaei , Somayeh Sojoudi
‹ 上一页 1 2 3 10 下一页 ›