中文
相关论文

相关论文: A hybrid deep learning method for finite-horizon m…

200 篇论文

We present a new framework to derandomise certain Markov chain Monte Carlo (MCMC) algorithms. As in MCMC, we first reduce counting problems to sampling from a sequence of marginal distributions. For the latter task, we introduce a method…

数据结构与算法 · 计算机科学 2023-04-05 Weiming Feng , Heng Guo , Chunyang Wang , Jiaheng Wang , Yitong Yin

We consider zero-sum stochastic games with finite state and action spaces, perfect information, mean payoff criteria, without any irreducibility assumption on the Markov chains associated to strategies (multichain games). The value of such…

最优化与控制 · 数学 2012-08-03 Marianne Akian , Jean Cochet-Terrasson , Sylvie Detournay , Stéphane Gaubert

Adaptive and interacting Markov chain Monte Carlo algorithms (MCMC) have been recently introduced in the literature. These novel simulation algorithms are designed to increase the simulation efficiency to sample complex distributions.…

统计理论 · 数学 2012-03-15 G. Fort , E. Moulines , P. Priouret

In the context of simple finite-state discrete time systems, we introduce a generalization of mean field game solution, called correlated solution, which can be seen as the mean field game analogue of a correlated equilibrium. Our notion of…

最优化与控制 · 数学 2021-07-12 Luciano Campi , Markus Fischer

In this paper we propose a high-order numerical scheme for time-dependent mean field games systems. The scheme, which is built by combining Lagrange-Galerkin and semi-Lagrangian techniques, is consistent and stable for large time steps…

数值分析 · 数学 2023-10-31 Elisa Calzola , Elisabetta Carlini , Francisco J. Silva

Variational quantum algorithms are poised to have significant impact on high-dimensional optimization, with applications in classical combinatorics, quantum chemistry, and condensed matter. Nevertheless, the optimization landscape of these…

量子物理 · 物理学 2022-02-02 Taylor L. Patti , Omar Shehab , Khadijeh Najafi , Susanne F. Yelin

Convergence rate analysis for general state-space Markov chains is fundamentally important in areas such as Markov chain Monte Carlo and algorithmic analysis (for computing explicit convergence bounds). This problem, however, is notoriously…

机器学习 · 计算机科学 2025-07-22 Yanlin Qu , Jose Blanchet , Peter Glynn

A stochastic incremental subgradient algorithm for the minimization of a sum of convex functions is introduced. The method sequentially uses partial subgradient information and the sequence of partial subgradients is determined by a general…

最优化与控制 · 数学 2021-08-24 Rafael Massambone , Eduardo F. Costa , Elias S. Helou

Markov chain Monte Carlo (MCMC) algorithms have become powerful tools for Bayesian inference. However, they do not scale well to large-data problems. Divide-and-conquer strategies, which split the data into batches and, for each batch, run…

统计计算 · 统计学 2017-07-18 Christopher Nemeth , Chris Sherlock

In this paper, we model one-day international cricket games as Markov processes, applying forward and inverse Reinforcement Learning (RL) to develop three novel tools for the game. First, we apply Monte-Carlo learning to fit a nonlinear…

机器学习 · 计算机科学 2021-03-09 Manohar Vohra , George S. D. Gordon

In this paper, we propose an initial value fomulation of the discrete mean field games on finite graphs (Graph MFG), and design a neural network based approach to solve it. Graph MFG describes infinite, non-cooperative and interactive…

数值分析 · 数学 2026-04-08 Yaxin Feng , Yang Xiang , Haomin Zhou

Many structured data-fitting applications require the solution of an optimization problem involving a sum over a potentially large number of measurements. Incremental gradient algorithms offer inexpensive iterations by sampling a subset of…

数值分析 · 计算机科学 2018-08-23 Michael P. Friedlander , Mark Schmidt

In this paper we consider the problem of computing the stationary distribution of nearly completely decomposable Markov processes, a well-established area in the classical theory of Markov processes with broad applications in the design,…

数值分析 · 数学 2025-06-19 Vasileios Kalantzis , Mark S. Squillante , Chai Wah Wu

We present a novel hybrid strategy based on machine learning to improve curvature estimation in the level-set method. The proposed inference system couples enhanced neural networks with standard numerical schemes to compute curvature more…

机器学习 · 计算机科学 2022-09-29 Luis Ángel Larios-Cárdenas , Frédéric Gibou

Markov chain Monte Carlo methods are a powerful tool for sampling equilibrium configurations in complex systems. One problem these methods often face is slow convergence over large energy barriers. In this work, we propose a novel method…

计算物理 · 物理学 2024-05-29 Luigi Sbailò , Manuel Dibak , Frank Noé

The goal of this paper is to analyze distributional Markov Decision Processes as a class of control problems in which the objective is to learn policies that steer the distribution of a cumulative reward toward a prescribed target law,…

最优化与控制 · 数学 2026-02-09 Nicole Bäuerle , Athanasios Vasileiadis

Objective: In a companion paper, we propose a parametric hybrid automaton model and an algorithm for the online synthesis of robustly correct and near-optimal controllers for cyber-physical system with reach-avoid guarantees. A key part of…

系统与控制 · 电气工程与系统科学 2025-02-10 Mario Gleirscher

We consider the simulation of Bayesian statistical inverse problems governed by large-scale linear and nonlinear partial differential equations (PDEs). Markov chain Monte Carlo (MCMC) algorithms are standard techniques to solve such…

数值分析 · 数学 2021-02-09 Harbir Antil , Howard C Elman , Akwum Onwunta , Deepanshu Verma

A mean-field method for the hypercubic nearest-neighbor Ising system is introduced and applications to the method are demonstrated. The main idea of this work is to combine the Kadanoff's mean-field approach with the model presented by one…

统计力学 · 物理学 2020-07-13 Tuncer Kaya , Başer Tambaş

Multi-agent reinforcement learning (MARL) is often modeled using the framework of Markov games (also called stochastic games or dynamic games). Most of the existing literature on MARL concentrates on zero-sum Markov games but is not…

计算机科学与博弈论 · 计算机科学 2022-12-20 Jayakumar Subramanian , Amit Sinha , Aditya Mahajan
‹ 上一页 1 8 9 10 下一页 ›