中文
相关论文

相关论文: Optimal Dynamic Information Provision

200 篇论文

An asymmetric information model is introduced for the situation in which there is a small agent who is more susceptible to the flow of information in the market than the general market participant, and who tries to implement strategies…

交易与市场微观结构 · 定量金融 2013-01-31 Dorje C. Brody , Mark H. A. Davis , Robyn L. Friedman , Lane P. Hughston

We consider the problem of designing an expected-revenue maximizing mechanism for allocating multiple non-perishable goods of $k$ varieties to flexible consumers over $T$ time steps. In our model, a random number of goods of each variety…

计算机科学与博弈论 · 计算机科学 2020-07-08 Shiva Navabi , Ashutosh Nayyar

Firms strategically disclose product information in order to attract consumers, but recipients often find it costly to process all of it, especially when products have complex features. We study a model of competitive information disclosure…

理论经济学 · 经济学 2020-02-04 Vasudha Jain , Mark Whitmeyer

In decision-making problems, the actions of an agent may reveal sensitive information that drives its decisions. For instance, a corporation's investment decisions may reveal its sensitive knowledge about market dynamics. To prevent this…

系统与控制 · 电气工程与系统科学 2020-04-17 Parham Gohari , Matthew Hale , Ufuk Topcu

Modelling all possible life cycles of a company in a highly competitive economic environment gives a significant advantage to the owner in his business investment activities. This article proposes and analyses a dynamic model of a company's…

综合金融 · 定量金融 2019-04-23 O. A. Malafeyev , I. I. Pavlov

The ability of a society to make the right decisions on relevant matters relies on its capability to properly aggregate the noisy information spread across the individuals it is made of. In this paper we study the information aggregation…

物理与社会 · 物理学 2015-06-16 Giacomo Livan , Matteo Marsili

In dynamic programming and reinforcement learning, the policy for the sequential decision making of an agent in a stochastic environment is usually determined by expressing the goal as a scalar reward function and seeking a policy that…

人工智能 · 计算机科学 2025-02-26 Simon Dima , Simon Fischer , Jobst Heitzig , Joss Oliver

Information-theoretic principles for learning and acting have been proposed to solve particular classes of Markov Decision Problems. Mathematically, such approaches are governed by a variational free energy principle and allow solving MDP…

人工智能 · 计算机科学 2016-04-08 Jordi Grau-Moya , Felix Leibfried , Tim Genewein , Daniel A. Braun

In online markets, agents often learn from other's actions in addition to their private information. Such observational learning can lead to herding or information cascades in which agents eventually ignore their private information and…

社会与信息网络 · 计算机科学 2025-05-16 Pawan Poojary , Randall Berry

We consider the disclosure problem of a sender with a large data set of hard evidence who wants to persuade a receiver to take higher actions. Because the receiver will make inferences based on the distribution of the data they see, the…

理论经济学 · 经济学 2023-11-03 Ying Gao

We study the design of autonomous agents that are capable of deceiving outside observers about their intentions while carrying out tasks in stochastic, complex environments. By modeling the agent's behavior as a Markov decision process, we…

人工智能 · 计算机科学 2021-09-15 Yagiz Savas , Christos K. Verginis , Ufuk Topcu

The hidden-action model captures a fundamental problem of principal-agent theory and provides an optimal sharing rule when only the outcome but not the effort can be observed. However, the hidden-action model builds on various explicit and…

综合经济学 · 经济学 2020-04-15 Stephan Leitner , Friederike Wall

Information-theoretic arguments focus on modeling the reliability of information transmission, assuming availability of infinite data at sources, thus ignoring randomness in message generation times at the respective sources. However, in…

网络与互联网体系结构 · 计算机科学 2009-09-29 K. C. V. Kalyanarama Sesha Sayee

In this paper, we explore a scenario where a sender provides an information policy and a receiver, upon observing a realization of this policy, decides whether to take a particular action, such as making a purchase. The sender's objective…

数值分析 · 数学 2024-12-13 Jorge Justiniano , Andreas Kleiner , Benny Moldovanu , Martin Rumpf , Philipp Strack

Most reinforcement learning methods are based upon the key assumption that the transition dynamics and reward functions are fixed, that is, the underlying Markov decision process is stationary. However, in many real-world applications, this…

机器学习 · 计算机科学 2020-09-23 Yash Chandak , Georgios Theocharous , Shiv Shankar , Martha White , Sridhar Mahadevan , Philip S. Thomas

We study a class of sequential decision-making problems with augmented predictions, potentially provided by a machine learning algorithm. In this setting, the decision-maker receives prediction intervals for unknown parameters that become…

机器学习 · 计算机科学 2025-05-05 Xin Chen , Yuze Chen , Yuan Zhou

We propose a dynamic spectrum access scheme where secondary users recommend "good" channels to each other and access accordingly. We formulate the problem as an average reward based Markov decision process. We show the existence of the…

分布式、并行与集群计算 · 计算机科学 2011-07-14 Xu Chen , Jianwei Huang , Husheng Li

General purpose intelligent learning agents cycle through (complex,non-MDP) sequences of observations, actions, and rewards. On the other hand, reinforcement learning is well-developed for small finite state Markov Decision Processes…

人工智能 · 计算机科学 2009-12-30 Marcus Hutter

The problem of making sequential decisions in unknown probabilistic environments is studied. In cycle $t$ action $y_t$ results in perception $x_t$ and reward $r_t$, where all quantities in general may depend on the complete history. The…

人工智能 · 计算机科学 2007-05-23 Marcus Hutter

We consider an intermediary's problem of dynamically matching demand and supply of heterogeneous types in a periodic-review fashion. More specifically, there are two disjoint sets of demand and supply types, and a reward associated with…

最优化与控制 · 数学 2018-11-20 Ming Hu , Yun Zhou