English
Related papers

Related papers: The Value Functions of Markov Decision Processes

200 papers

Productions functions map the inputs of a firm or a productive system onto its outputs. This article expounds generalizations of the production function that include state variables, organizational structures and increasing returns to…

Physics and Society · Physics 2008-12-02 Guido Fioretti

Markov decision processes (MDPs) are used to model a wide variety of applications ranging from game playing over robotics to finance. Their optimal policy typically maximizes the expected sum of rewards given at each step of the decision…

Machine Learning · Computer Science 2025-05-26 Maximilian Nägele , Jan Olle , Thomas Fösel , Remmy Zen , Florian Marquardt

In this paper we propose a general approach to define a many-valued preferential interpretation of gradual argumentation semantics. The approach allows for conditional reasoning over arguments and boolean combination of arguments, with…

Artificial Intelligence · Computer Science 2025-06-10 Mario Alviano , Laura Giordano , Daniele Theseider Dupré

Motivated by existing results, we present some completely monotonic functions involving the polygamma functions.

Classical Analysis and ODEs · Mathematics 2010-12-03 Peng Gao

We introduce a new approach for deterministic sensitivity analysis of Markov reward processes, commonly used in cost-effectiveness analyses, via reformulation into a polynomial system. Our approach leverages cylindrical algebraic…

Optimization and Control · Mathematics 2024-10-10 Timothy C. Y. Chan , Muhammad Maaz

The online Markov decision process (MDP) is a generalization of the classical Markov decision process that incorporates changing reward functions. In this paper, we propose practical online MDP algorithms with policy iteration and…

Machine Learning · Computer Science 2015-10-16 Yao Ma , Hao Zhang , Masashi Sugiyama

We prove martingale-ergodic and ergodic-martingale theorems for vector valued Bochner integrable functions. We obtain dominant and maximal inequalities. We also prove weighted and multiparameter martingale-ergodic and ergodic martingale…

Functional Analysis · Mathematics 2012-01-10 Farruh Shahidi , Inomjon Ganiev

The Kolmogorov axioms for probability functions are placed in the context of signed meadows. A completeness theorem is stated and proven for the resulting equational theory of probability calculus. Elementary definitions of probability…

Logic · Mathematics 2016-12-23 Jan A. Bergstra , Alban Ponse

We consider the challenge of preference elicitation in systems that help users discover the most desirable item(s) within a given database. Past work on preference elicitation focused on structured models that provide a factored…

Artificial Intelligence · Computer Science 2012-07-19 Ronen I. Brafman , Carmel Domshlak , Tanya Kogan

In this paper the set of value functions of all-possible zero-sum differential games with terminal payoff is characterized. The necessary and sufficient condition for a given function to be a value of some differential game with terminal…

Optimization and Control · Mathematics 2008-11-12 Yurii Averboukh

We aim to link random fields and marked point processes and therefore introduce a new class of stochastic processes which are defined on a random set in R^d. Unlike for random fields, the mark covariance function of a marked random set is…

Probability · Mathematics 2012-01-25 Felix Ballani , Zakhar Kabluchko , Martin Schlather

Value decomposition has long been a fundamental technique in multi-agent dynamic programming and reinforcement learning (RL). Specifically, the value function of a global state $(s_1,s_2,\ldots,s_N)$ is often approximated as the sum of…

Machine Learning · Computer Science 2025-11-14 Shuze Chen , Tianyi Peng

In classic reinforcement learning (RL) and decision making problems, policies are evaluated with respect to a scalar reward function, and all optimal policies are the same with regards to their expected return. However, many real-world…

Machine Learning · Computer Science 2023-11-02 Han Shao , Lee Cohen , Avrim Blum , Yishay Mansour , Aadirupa Saha , Matthew R. Walter

In this paper, a new kind of soft sets related with some common decision making problems in real life called central soft sets is introduced. Properties of some basic operations on central soft sets are shown. It is investigated that some…

Logic in Computer Science · Computer Science 2015-06-10 Xuechong Guan

We describe an approach for exploiting structure in Markov Decision Processes with continuous state variables. At each step of the dynamic programming, the state space is dynamically partitioned into regions where the value function is the…

Artificial Intelligence · Computer Science 2012-07-19 Zhengzhu Feng , Richard Dearden , Nicolas Meuleau , Richard Washington

Given a batch of human computation tasks, a commonly ignored aspect is how the price (i.e., the reward paid to human workers) of these tasks must be set or varied in order to meet latency or cost constraints. Often, the price is set…

Computer Science and Game Theory · Computer Science 2014-08-28 Yihan Gao , Aditya Parameswaran

We present a list of algebraic, combinatorial, and analytic mechanisms that give rise to determinantal point processes.

Probability · Mathematics 2009-11-09 Alexei Borodin

We formulate a probabilistic Markov property in discrete time under a dynamic risk framework with minimal assumptions. This is useful for recursive solutions to risk-sensitive versions of dynamic optimisation problems such as optimal…

Optimization and Control · Mathematics 2022-09-05 Tomasz Kosmala , Randall Martyr , John Moriarty

Valuations, as additive functionals, allow various applications in Stochastic Geometry, yielding mean value formulas for specific random closed sets and processes of convex or polyconvex particles. In particular, valuations are especially…

Probability · Mathematics 2015-10-28 Julia Hörrmann , Wolfgang Weil

Markov decision processes (MDPs) are a popular model for performance analysis and optimization of stochastic systems. The parameters of stochastic behavior of MDPs are estimates from empirical observations of a system; their values are not…

Artificial Intelligence · Computer Science 2017-10-26 Dimitri Scheftelowitsch , Peter Buchholz , Vahid Hashemi , Holger Hermanns
‹ Prev 1 8 9 10 Next ›