English
Related papers

Related papers: A note on the article "On Exploiting Spectral Prop…

200 papers

Reinforcement learning (RL) in Markov decision processes (MDPs) with large state spaces is a challenging problem. The performance of standard RL algorithms degrades drastically with the dimensionality of state space. However, in practice,…

Artificial Intelligence · Computer Science 2018-06-21 Kamyar Azizzadenesheli , Alessandro Lazaric , Animashree Anandkumar

We consider the problem of average consensus in a distributed system comprising a set of nodes that can exchange information among themselves. We focus on a class of algorithms for solving such a problem whereby each node maintains a state…

Multiagent Systems · Computer Science 2024-03-12 Christoforos N. Hadjicostis , Alejandro D. Dominguez-Garcia

We study an approximation method for partially observed Markov decision processes (POMDPs) with continuous spaces. Belief MDP reduction, which has been the standard approach to study POMDPs requires rigorous approximation methods for…

Optimization and Control · Mathematics 2025-01-20 Ali Devran Kara , Erhan Bayraktar , Serdar Yuksel

We study the effect of strategic behavior in iterative voting for multiple issues under uncertainty. We introduce a model synthesizing simultaneous multi-issue voting with Meir, Lev, and Rosenschein (2014)'s local dominance theory and…

Computer Science and Game Theory · Computer Science 2023-01-24 Joshua Kavner , Reshef Meir , Francesca Rossi , Lirong Xia

Raghavendra (STOC 2008) gave an elegant and surprising result: if Khot's Unique Games Conjecture (STOC 2002) is true, then for every constraint satisfaction problem (CSP), the best approximation ratio is attained by a certain simple…

Data Structures and Algorithms · Computer Science 2010-11-01 Yuichi Yoshida

We study convergence properties of pseudo-marginal Markov chain Monte Carlo algorithms (Andrieu and Roberts [Ann. Statist. 37 (2009) 697-725]). We find that the asymptotic variance of the pseudo-marginal algorithm is always at least as…

Probability · Mathematics 2015-03-31 Christophe Andrieu , Matti Vihola

We propose a semidefinite programming (SDP) algorithm for community detection in the stochastic block model, a popular model for networks with latent community structure. We prove that our algorithm achieves exact recovery of the latent…

Data Structures and Algorithms · Computer Science 2016-12-05 Amelia Perry , Alexander S. Wein

We study maximum-entropy inference for finite-dimensional quantum states under linear moment constraints. Given expectation values of finitely many observables, the feasible set of states is convex but typically non-unique. The…

Quantum Physics · Physics 2025-10-27 James Tian

Standard Markov decision process (MDP) and reinforcement learning algorithms optimize the policy with respect to the expected gain. We propose an algorithm which enables to optimize an alternative objective: the probability that the gain is…

Machine Learning · Computer Science 2023-03-06 Vincent Corlay , Jean-Christophe Sibel

Landmarks$\unicode{x2013}$conditions that must be satisfied at some point in every solution plan$\unicode{x2013}$have contributed to major advancements in classical planning, but they have seldom been used in stochastic domains. We…

Artificial Intelligence · Computer Science 2025-08-18 David H. Chan , Mark Roberts , Dana S. Nau

Sufficient conditions are identified under which the value function and the optimal strategy of a Markov decision process (MDP) are even and quasi-convex in the state. The key idea behind these conditions is the following. First, sufficient…

Optimization and Control · Mathematics 2017-09-12 Jhelum Chakravorty , Aditya Mahajan

The Lasserre hierarchy of semidefinite programming (SDP) relaxations is an effective scheme for finding computationally feasible SDP approximations of polynomial optimization over compact semi-algebraic sets. In this paper, we show that,…

Optimization and Control · Mathematics 2013-06-28 V. Jeyakumar , T. S. Pham , G. Li

The online Markov decision process (MDP) is a generalization of the classical Markov decision process that incorporates changing reward functions. In this paper, we propose practical online MDP algorithms with policy iteration and…

Machine Learning · Computer Science 2015-10-16 Yao Ma , Hao Zhang , Masashi Sugiyama

The spectral gap problem - determining whether the energy spectrum of a system has an energy gap above ground state, or if there is a continuous range of low-energy excitations - pervades quantum many-body physics. Recently, this important…

Quantum Physics · Physics 2020-08-19 Johannes Bausch , Toby Cubitt , Angelo Lucia , David Perez-Garcia

Partial differential equations with discrete (concentrated) state-dependent delays are studied. The existence and uniqueness of solutions with initial data from a wider linear space is proven first and then a subset of the space of…

Analysis of PDEs · Mathematics 2010-11-11 Alexander V. Rezounenko , Petr Zagalak

We build on a recently introduced geometric interpretation of Markov Decision Processes (MDPs) to analyze classical MDP-solving algorithms: Value Iteration (VI) and Policy Iteration (PI). First, we develop a geometry-based analytical…

Machine Learning · Computer Science 2025-03-07 Arsenii Mustafin , Aleksei Pakharev , Alex Olshevsky , Ioannis Ch. Paschalidis

This paper is concerned with the sample efficiency of reinforcement learning, assuming access to a generative model (or simulator). We first consider $\gamma$-discounted infinite-horizon Markov decision processes (MDPs) with state space…

Machine Learning · Computer Science 2025-03-18 Gen Li , Yuting Wei , Yuejie Chi , Yuxin Chen

We study entanglement-related properties of random quantum states which are unitarily invariant, in the sense that their distribution is left unchanged by conjugation with arbitrary unitary operators. In the large matrix size limit, the…

Mathematical Physics · Physics 2018-07-09 Ion Nechita

In this work we are interested in stochastic particle methods for multi-objective optimization. The problem is formulated using parametrized, single-objective sub-problems which are solved simultaneously. To this end a consensus based…

Optimization and Control · Mathematics 2022-08-03 Giacomo Borghi , Michael Herty , Lorenzo Pareschi

Perfectly Matched Layers (PML) has become a very common method for the numerical approximation of wave and wave-like equations on unbounded domains. This technique allows one to obtain accurate solutions while working on a finite…

Analysis of PDEs · Mathematics 2025-03-11 Kurt Bryan , Michael S. Vogelius