中文
相关论文

相关论文: A hybrid deep learning method for finite-horizon m…

200 篇论文

Hierarchical Reinforcement Learning (HRL) approaches have shown successful results in solving a large variety of complex, structured, long-horizon problems. Nevertheless, a full theoretical understanding of this empirical evidence is…

机器学习 · 计算机科学 2025-02-05 Gianluca Drappo , Alberto Maria Metelli , Marcello Restelli

Hybrid systems, and Piecewise Deterministic Markov Processes in particular, are widely used to model and numerically study systems exhibiting multiple time scales in biochemical reaction kinetics and related areas. In this paper an almost…

数值分析 · 数学 2011-12-07 Martin G. Riedler

This paper focuses on discussing Newton's method and its hybrid with machine learning for the steady state Navier-Stokes Darcy model discretized by mixed element methods. First, a Newton iterative method is introduced for solving the…

数值分析 · 数学 2024-03-07 Jianguo Huang , Hui Peng , Haohao Wu

Recently, Sidford, Wang, Wu and Ye (2018) developed an algorithm combining variance reduction techniques with value iteration to solve discounted Markov decision processes. This algorithm has a sublinear complexity when the discount factor…

最优化与控制 · 数学 2019-09-16 Marianne Akian , Stéphane Gaubert , Zheng Qu , Omar Saadi

We develop the fictitious play algorithm in the context of the linear programming approach for mean field games of optimal stopping and mean field games with regular control and absorption. This algorithm allows to approximate the mean…

最优化与控制 · 数学 2023-01-25 Roxana Dumitrescu , Marcos Leutscher , Peter Tankov

We propose a data-driven mean-curvature solver for the level-set method. This work is the natural extension to $\mathbb{R}^3$ of our two-dimensional strategy in [DOI: 10.1007/s10915-022-01952-2][1] and the hybrid inference system of [DOI:…

机器学习 · 计算机科学 2022-12-12 Luis Ángel Larios-Cárdenas , Frédéric Gibou

We introduce a new Markov Chain Monte Carlo (MCMC) algorithm with parallel tempering for fitting theoretical models of horizon-scale images of black holes to the interferometric data from the Event Horizon Telescope (EHT). The algorithm…

Markov Chain Monte Carlo methods are widely used in signal processing and communications for statistical inference and stochastic optimization. In this work, we introduce an efficient adaptive Metropolis-Hastings algorithm to draw samples…

统计计算 · 统计学 2016-03-17 David Luengo , Luca Martino

Multi-agent reinforcement learning methods have shown remarkable potential in solving complex multi-agent problems but mostly lack theoretical guarantees. Recently, mean field control and mean field games have been established as a…

机器学习 · 计算机科学 2021-12-20 Kai Cui , Anam Tahir , Mark Sinzger , Heinz Koeppl

We demonstrate the use of a variational method to determine a quantitative lower bound on the rate of convergence of Markov Chain Monte Carlo (MCMC) algorithms as a function of the target density and proposal density. The bound relies on…

数据分析、统计与概率 · 物理学 2013-05-29 Fergal P. Casey , Joshua J. Waterfall , Ryan N. Gutenkunst , Christopher R. Myers , James P. Sethna

In this paper, we apply the Monte Carlo stochastic optimization (MOST) proposed by the authors to a deep learning of XOR gate and verify its effectiveness. Deep machine learning based on neural networks is one of the most important keywords…

机器学习 · 计算机科学 2021-09-07 Sin-ichi Inage , Hana Hebishima

We consider discrete-time stationary mean field games (MFG) with unknown dynamics and design algorithms for finding the equilibrium with finite-time complexity guarantees. Prior solutions to the problem assume either the contraction of a…

最优化与控制 · 数学 2025-02-13 Sihan Zeng , Sujay Bhatt , Alec Koppel , Sumitra Ganesh

We introduce a gradient-based learning method to automatically adapt Markov chain Monte Carlo (MCMC) proposal distributions to intractable targets. We define a maximum entropy regularised objective function, referred to as generalised speed…

机器学习 · 统计学 2020-01-07 Michalis K. Titsias , Petros Dellaportas

We present a highly efficient proximal Markov chain Monte Carlo methodology to perform Bayesian computation in imaging problems. Similarly to previous proximal Monte Carlo approaches, the proposed method is derived from an approximation of…

统计计算 · 统计学 2020-03-20 Luis Vargas , Marcelo Pereyra , Konstantinos C. Zygalakis

In the context of Markov decision processes running in continuous time, one of the most intriguing challenges is the efficient approximation of finite horizon reachability objectives. A multitude of sophisticated model checking algorithms…

系统与控制 · 计算机科学 2015-08-03 Yuliya Butkova , Hassan Hatefi , Holger Hermanns , Jan Krcal

An algorithm for estimating quasi-stationary distribution of finite state space Markov chains has been proven in a previous paper. Now this paper proves a similar algorithm that works for general state space Markov chains under very general…

概率论 · 数学 2015-03-04 Jose H. Blanchet , Peter Glynn , Shuheng Zheng

We propose a simple and original approach for solving linear-quadratic mean-field stochastic control problems. We study both finite-horizon and infinite-horizon problems, and allow notably some coefficients to be stochastic. Our method is…

概率论 · 数学 2017-11-28 Matteo Basei , Huyên Pham

A class of nonzero-sum stochastic dynamic games with imperfect information structure is investigated. The game involves an arbitrary number of players, modeled as homogeneous Markov decision processes, aiming to find a sequential Nash…

最优化与控制 · 数学 2019-12-17 Jalal Arabneydi , Amir G. Aghdam

In this paper we propose new algorithm to reduce autocorrelation in Markov chain Monte-Carlo algorithms for euclidean field theories on the lattice. Our proposing algorithm is the Hybrid Monte-Carlo algorithm (HMC) with restricted Boltzmann…

高能物理 - 格点 · 物理学 2017-12-13 Akinori Tanaka , Akio Tomiya

This paper proposes and analyzes two neural network methods to solve the master equation for finite-state mean field games (MFGs). Solving MFGs provides approximate Nash equilibria for stochastic, differential games with finite but large…

最优化与控制 · 数学 2024-12-24 Asaf Cohen , Mathieu Laurière , Ethan Zell