中文
相关论文

相关论文: Discounting in Strategy Logic

200 篇论文

It has been recently pointed out that dynamical systems depending on future values of the unknowns may be useful in different areas of knowledge. We explore in this context the extension of the concept of order reduction that has been…

计算物理 · 物理学 2007-05-23 J. M. Aguirregabiria

The appropriate discount rate for evaluating policies is a critical consideration in economic decision-making. This paper presents a new model for calculating the derived discount rate for a society that includes different groups with…

理论经济学 · 经济学 2025-02-11 Mahdi Mousavi , Mahdi Kohan Sefidi

A prominent theme in behavioural contract theory is the study of present-biased agents represented through quasi-hyperbolic discounting. In a model of competitive credit provision, we study an alternative to this framework in which the…

理论经济学 · 经济学 2026-02-11 Siddharth Chatterjee , Daniel F. Garrett

In this paper, we present a model of Partnership Game with respect to the important role of partnership and cooperation in nowdays life. Since such interactions are repeated frequently, we study this model as a Stage Game in the structure…

动力系统 · 数学 2019-02-18 E. Sorouri , M. Eshaghi Gordji

Learning in sparse reward settings remains a challenge in Reinforcement Learning, which is often addressed by using intrinsic rewards. One promising strategy is inspired by human curiosity, requiring the agent to learn to predict the…

机器学习 · 计算机科学 2018-10-02 Gino Brunner , Manuel Fritsche , Oliver Richter , Roger Wattenhofer

Inspired by real-time ad exchanges for online display advertising, we consider the problem of inferring a buyer's value distribution for a good when the buyer is repeatedly interacting with a seller through a posted-price mechanism. We…

机器学习 · 计算机科学 2013-11-28 Kareem Amin , Afshin Rostamizadeh , Umar Syed

The model checking problem for multi-agent systems against Strategy Logic specifications is known to be non-elementary. On this logic several fragments have been defined to tackle this issue but at the expense of expressiveness. In this…

多智能体系统 · 计算机科学 2023-10-27 Francesco Belardinelli , Angelo Ferrando , Wojciech Jamroga , Vadim Malvone , Aniello Murano

We revisit multi-agent asynchronous online optimization with delays, where only one of the agents becomes active for making the decision at each round, and the corresponding feedback is received by all the agents after unknown delays.…

机器学习 · 计算机科学 2026-02-05 Lingchan Bao , Tong Wei , Yuanyu Wan

The development of logic has largely been through the 'deductive' paradigm: conclusions are inferred from established premisses. However, the use of logic in the context of both human and machine reasoning is typically through the dual…

计算机科学中的逻辑 · 计算机科学 2025-04-29 Alexander V. Gheorghiu , David J. Pym

This paper is concerned with the determination of pricing strategies for a firm that in each period of a finite horizon receives replenishment quantities of a single product which it sells in two markets, e.g., a long-distance market and an…

最优化与控制 · 数学 2015-09-25 Wen , Chen , Adam Fleischhacker , Michael N. Katehakis

Complexity theory is a useful tool to study computational issues surrounding the elicitation of preferences, as well as the strategic manipulation of elections aggregating together preferences of multiple agents. We study here the…

人工智能 · 计算机科学 2012-04-18 Toby Walsh

Strategy Logic (SL, for short) has been recently introduced by Mogavero, Murano, and Vardi as a useful formalism for reasoning explicitly about strategies, as first-order objects, in multi-agent concurrent games. This logic turns to be very…

计算机科学中的逻辑 · 计算机科学 2015-03-20 Fabio Mogavero , Aniello Murano , Giuseppe Perelli , Moshe Y. Vardi

Evaluation metrics are an essential part of a ranking system, and in the past many evaluation metrics have been proposed in information retrieval and Web search. Discounted Cumulated Gains (DCG) has emerged as one of the evaluation metrics…

信息检索 · 计算机科学 2012-12-27 Ke Zhou , Hongyuan Zha , Gui-Rong Xue , Yong Yu

This note provides upper bounds on the number of operations required to compute by value iterations a nearly optimal policy for an infinite-horizon discounted Markov decision process with a finite number of states and actions. For a given…

最优化与控制 · 数学 2020-01-29 Eugene A. Feinberg , Gaojin He

Linear temporal logic (LTL) is a powerful language for task specification in reinforcement learning, as it allows describing objectives beyond the expressivity of conventional discounted return formulations. Nonetheless, recent works have…

机器学习 · 计算机科学 2025-06-11 Marco Bagatella , Andreas Krause , Georg Martius

While agent evaluation has shifted toward long-horizon tasks, most benchmarks still emphasize local, step-level reasoning rather than the global constrained optimization (e.g., time and financial budgets) that demands genuine planning…

人工智能 · 计算机科学 2026-01-27 Yinger Zhang , Shutong Jiang , Renhao Li , Jianhong Tu , Yang Su , Lianghao Deng , Xudong Guo , Chenxu Lv , Junyang Lin

We present a method to find an optimal policy with respect to a reward function for a discounted Markov decision process under general linear temporal logic (LTL) specifications. Previous work has either focused on maximizing a cumulative…

系统与控制 · 电气工程与系统科学 2021-03-24 Krishna C. Kalagarla , Rahul Jain , Pierluigi Nuzzo

In an effort to better understand the different ways in which the discount factor affects the optimization process in reinforcement learning, we designed a set of experiments to study each effect in isolation. Our analysis reveals that the…

机器学习 · 计算机科学 2019-12-24 Harm van Seijen , Mehdi Fatemi , Arash Tavakoli

An approximation of strategyproofness in large, two-sided matching markets is highly evident. Through simulations, one can observe that the percentage of agents with useful deviations decreases as the market size grows. Furthermore, there…

多智能体系统 · 计算机科学 2022-11-30 Lars Lien Ankile , Kjartan Krange , Yuto Yagi

Additively separable hedonic games and fractional hedonic games have received considerable attention. They are coalition forming games of selfish agents based on their mutual preferences. Most of the work in the literature characterizes the…

人工智能 · 计算机科学 2017-06-29 Michele Flammini , Gianpiero Monaco , Qiang Zhang