中文
相关论文

相关论文: Optimal Dynamic Information Provision

200 篇论文

We study offline reinforcement learning problems with a long-run average reward objective. The state-action pairs generated by any fixed behavioral policy thus follow a Markov chain, and the {\em empirical} state-action-next-state…

最优化与控制 · 数学 2025-03-18 Mengmeng Li , Daniel Kuhn , Tobias Sutter

An uninformed sender publicly commits to an informative experiment about an uncertain state, privately observes its outcome, and sends a cheap-talk message to a receiver. We provide an algorithm valid for arbitrary state-dependent…

理论经济学 · 经济学 2024-05-10 Qianjun Lyu , Wing Suen

We study how queue-state information disclosures affect impatient tenants in multi-tenant edge systems. We propose an information-bulletin strategy in which each queue periodically broadcasts two Markov models. One is a model of…

系统与控制 · 电气工程与系统科学 2025-10-15 Anthony Kiggundu , Bin Han , Hans D. Schotten

The distributionally robust Markov Decision Process (MDP) approach asks for a distributionally robust policy that achieves the maximal expected total reward under the most adversarial distribution of uncertain parameters. In this paper, we…

系统与控制 · 计算机科学 2018-10-10 Zhi Chen , Pengqian Yu , William B. Haskell

In today's economy, it becomes important for Internet platforms to consider the sequential information design problem to align its long term interest with incentives of the gig service providers. This paper proposes a novel model of…

人工智能 · 计算机科学 2022-02-23 Jibang Wu , Zixuan Zhang , Zhe Feng , Zhaoran Wang , Zhuoran Yang , Michael I. Jordan , Haifeng Xu

An agent acquires information dynamically until her belief about a binary state reaches an upper or lower threshold. She can choose any signal process subject to a constraint on the rate of entropy reduction. Strategies are ordered by "time…

理论经济学 · 经济学 2024-08-23 Daniel Chen , Weijie Zhong

We consider a class of distributed submodular maximization problems in which each agent must choose a single strategy from its strategy set. The global objective is to maximize a submodular function of the strategies chosen by each agent.…

数据结构与算法 · 计算机科学 2017-06-14 Bahman Gharesifard , Stephen L. Smith

We study a dynamic model of Bayesian persuasion in sequential decision-making settings. An informed principal observes an external parameter of the world and advises an uninformed agent about actions to take over time. The agent takes…

计算机科学与博弈论 · 计算机科学 2022-05-25 Jiarui Gan , Rupak Majumdar , Goran Radanovic , Adish Singla

We consider a model where an agent has a repeated decision to make and wishes to maximize their total payoff. Payoffs are influenced by an action taken by the agent, but also an unknown state of the world that evolves over time. Before…

计算机科学与博弈论 · 计算机科学 2021-01-20 Nicole Immorlica , Ian Kash , Brendan Lucier

Consider a dynamic decision-making scenario where at every stage the investor has to choose between investing in one of two projects or gathering more information. At each stage, the investor may seek counsel from one of several…

信息论 · 计算机科学 2024-05-01 Yuval Cornfeld , Ehud Lehrer , Eilon Solan

This paper considers a half-duplex scenario where an interferer behaves according to a parametric model but the values of the model parameters are unknown. We explore the necessary number of sensing steps to gather sufficient knowledge…

信息论 · 计算机科学 2024-10-11 Vincent Corlay , Jean-Christophe Sibel , Nicolas Gresset

We develop a Bayesian model for decision-making under time pressure with endogenous information acquisition. In our model, the decision maker decides when to observe (costly) information by sampling an underlying continuous-time stochastic…

人工智能 · 计算机科学 2016-10-25 Ahmed M. Alaa , Mihaela van der Schaar

A decision maker typically (i) incorporates training data to learn about the relative effectiveness of treatments, and (ii) chooses an implementation mechanism that implies an ``optimal'' predicted outcome distribution according to some…

计量经济学 · 经济学 2025-05-29 Anders Bredahl Kock , David Preinerstorfer

A key issue in the control of distributed discrete systems modeled as Markov decisions processes, is that often the state of the system is not directly observable at any single location in the system. The participants in the control scheme…

信息论 · 计算机科学 2017-05-01 Jie Ren , Solmaz Torabi , John MacLaren Walsh

Neglecting the effect that decisions have on individuals (and thus, on the underlying data distribution) when designing algorithmic decision-making policies may increase inequalities and unfairness in the long term - even if fairness…

人工智能 · 计算机科学 2023-11-22 Miriam Rateike , Isabel Valera , Patrick Forré

We consider a model of a data broker selling information to a single agent to maximize his revenue. The agent has a private valuation of the additional information, and upon receiving the signal from the data broker, the agent can conduct…

理论经济学 · 经济学 2023-08-08 Yingkai Li

We study multi-product monopoly pricing where the seller jointly designs the selling mechanism and the information structure for the buyer to learn his values. Unlike the case with exogenous information, we show that when the seller…

计算机科学与博弈论 · 计算机科学 2025-10-29 Yang Cai , Yingkai Li , Jinzhao Wu

We consider stopping problems in which a decision maker (DM) faces an unknown state of nature and decides sequentially whether to stop and take an irreversible action; pay a fee and obtain additional information; or wait without acquiring…

理论经济学 · 经济学 2022-05-16 Ehud Lehrer , Tao Wang

Information relaxation and duality in Markov decision processes have been studied recently by several researchers with the goal to derive dual bounds on the value function. In this paper we extend this dual formulation to controlled Markov…

最优化与控制 · 数学 2014-10-23 Fan Ye , Enlu Zhou

We study a dynamic game where an expert sends probabilistic forecasts to a decision-maker. The decision-maker verifies these forecasts using a calibration test based on past data. How should the expert send forecasts to maximize her payoff…

理论经济学 · 经济学 2026-05-13 Atulya Jain , Vianney Perchet