中文
相关论文

相关论文: Rational hyperbolic discounting

200 篇论文

A set of objects is to be divided fairly among agents with different tastes, modeled by additive utility-functions. If we consider the objects as indivisible, many instances of the decision problem: ``Is there a fair division of the objects…

计算机科学与博弈论 · 计算机科学 2025-07-03 Samuel Bismuth , Ivan Bliznets , Erel Segal-Halevi

In reinforcement learning (RL), the goal is to obtain an optimal policy, for which the optimality criterion is fundamentally important. Two major optimality criteria are average and discounted rewards. While the latter is more popular, it…

机器学习 · 计算机科学 2022-09-05 Vektor Dewanto , Marcus Gallagher

Our goal is for agents to optimize the right reward function, despite how difficult it is for us to specify what that is. Inverse Reinforcement Learning (IRL) enables us to infer reward functions from demonstrations, but it usually assumes…

机器学习 · 计算机科学 2019-06-25 Rohin Shah , Noah Gundotra , Pieter Abbeel , Anca D. Dragan

Fair division has long been an important problem in the economics literature. In this note, we consider the existence of proportionally fair allocations of indivisible goods, i.e., allocations of indivisible goods in which every agent gets…

计算机科学与博弈论 · 计算机科学 2019-08-19 Warut Suksompong

This paper considers an infinite-horizon Markov decision process (MDP) that allows for general non-exponential discount functions, in both discrete and continuous time. Due to the inherent time inconsistency, we look for a randomized…

最优化与控制 · 数学 2024-12-10 Erhan Bayraktar , Yu-Jui Huang , Zhenhua Wang , Zhou Zhou

Inflation is painful, for firms, customers, employees, and society. But careful study of periods of hyperinflation point to ways that firms can adapt. In particular, companies need to think about how to change prices regularly and cheaply,…

综合经济学 · 经济学 2022-12-02 Mark Bergen , Thomas Bergen , Daniel Levy , Rose Semenov

An ability to postpone one's execution without penalty provides an important strategic advantage in high-frequency trading. To elucidate competition between traders one has to formulate to a quantitative theory of formation of the execution…

交易与市场微观结构 · 定量金融 2014-06-20 Peter Lerner

To determine the welfare implications of price changes in demand data, we introduce a revealed preference relation over prices. We show that the absence of cycles in this relation characterizes a consumer who trades off the utility of…

计量经济学 · 经济学 2024-07-03 Rahul Deb , Yuichi Kitamura , John K. -H. Quah , Jörg Stoye

We review the theory of renewal reward processes, which describes renewal processes that have some cost or reward associated with each cycle. We present a new simplified proof of the renewal reward theorem that mimics the proof of the…

概率论 · 数学 2014-04-23 Maria Vlasiou

Consider a game where Alice generates an integer and Bob wins if he can factor that integer. Traditional game theory tells us that Bob will always win this game even though in practice Alice will win given our usual assumptions about the…

计算机科学与博弈论 · 计算机科学 2009-11-18 Lance Fortnow , Rahul Santhanam

One method to offer some bidders a discount in a first-price auction is to augment their bids when selecting a winner but only charge them their original bids should they win. Another method is to use their original bids to select a winner,…

计算机科学与博弈论 · 计算机科学 2024-03-12 Miguel Alcobendas , Eric Bax

I provide a model of rational inattention with heterogeneity and prove it is observationally equivalent to a state-dependent stochastic choice model subject to attention costs. I demonstrate that additive separability of unobservable…

计量经济学 · 经济学 2024-02-16 Martin Bustos

Incentives have surprisingly inconsistent effects when it comes to encouraging people to behave prosocially. Classical economic theory, according to which a specific behavior becomes more prevalent when it is rewarded, struggles to explain…

综合经济学 · 经济学 2021-04-29 Caroline Graf , Eva-Maria Merz , Bianca Suanet , Pamala Wiepking

In repeated games, cooperation is possible in equilibrium only if players are sufficiently patient, and long-term gains from cooperation outweigh short-term gains from deviation. What happens if the players have incomplete information…

经济学 · 定量金融 2019-01-23 Cy Maor , Eilon Solan

The consumption function maps current wealth and the exogenous state to current consumption. We prove the existence and uniqueness of a consumption function when the agent has a preference for wealth. When the period utility functions are…

理论经济学 · 经济学 2025-09-30 Qingyin Ma , Alexis Akira Toda

Generally accepted depreciation methods do not compute the intrinsic value of an asset, as they do not factor for the Time Value of Money, a key principle within financial theory. This is disadvantageous, as knowing the intrinsic value of…

综合金融 · 定量金融 2018-01-22 Brendon Farrell

Problem definition: Mining for heterogeneous responses to an intervention is a crucial step for data-driven operations, for instance to personalize treatment or pricing. We investigate how to estimate price sensitivity from…

统计方法学 · 统计学 2025-01-08 Jean Pauphilet

LLMs utilizing chain-of-thought reasoning often waste substantial compute by producing long, incorrect responses. Abstention can mitigate this by withholding outputs unlikely to be correct. While most abstention methods decide to withhold…

The absent-minded driver's problem illustrates that probabilistic strategies can give higher pay-offs than deterministic ones. We show that there are strategies using quantum entangled states that give even higher pay-offs, both for the…

量子物理 · 物理学 2009-07-28 Adan Cabello , John Calsamiglia

We present a model of optimal training of a rational, sluggish agent. A trainer commits to a discrete-time, finite-state Markov process that governs the evolution of training intensity. Subsequently, the agent monitors the state and adjusts…

理论经济学 · 经济学 2021-05-20 Kfir Eliaz , Ran Spiegler
‹ 上一页 1 8 9 10 下一页 ›