中文
相关论文

相关论文: A hybrid deep learning method for finite-horizon m…

200 篇论文

In this paper, we develop a Mean Field Games approach to Cluster Analysis. We consider a finite mixture model, given by a convex combination of probability density functions, to describe the given data set. We interpret a data point as an…

数值分析 · 数学 2019-12-24 Laura Aquilanti , Simone Cacace , Fabio Camilli , Raul De Maio

An interesting iterative procedure is proposed to solve a two-player zero-sum Markov games. Under suitable assumption, the boundedness of the proposed iterates is obtained theoretically. Using results from stochastic approximation, the…

机器学习 · 计算机科学 2025-09-23 Shreyas S R , Antony Vijesh

This paper considers two-player zero-sum finite-horizon Markov games with simultaneous moves. The study focuses on the challenging settings where the value function or the model is parameterized by general function classes. Provably…

计算机科学与博弈论 · 计算机科学 2021-11-02 Baihe Huang , Jason D. Lee , Zhaoran Wang , Zhuoran Yang

Mean field games (MFGs) offer a powerful framework for modeling large-scale multi-agent systems. This paper addresses MFGs formulated in continuous time with discrete state spaces, where agents' dynamics are governed by continuous-time…

计算机科学与博弈论 · 计算机科学 2026-02-27 Yannick Eich , Christian Fabian , Kai Cui , Heinz Koeppl

Recent techniques have been successful in reconstructing surfaces as level sets of learned functions (such as signed distance fields) parameterized by deep neural networks. Many of these methods, however, learn only closed surfaces and are…

计算机视觉与模式识别 · 计算机科学 2022-03-23 David Palmer , Dmitriy Smirnov , Stephanie Wang , Albert Chern , Justin Solomon

We present a novel and mathematically transparent approach to function approximation and the training of large, high-dimensional neural networks, based on the approximate least-squares solution of associated Fredholm integral equations of…

数值分析 · 数学 2024-07-17 Patrick Gelß , Aizhan Issagali , Ralf Kornhuber

The covariance matrix of measurements of Markov random fields (processes) has useful properties that allow to develop effective computational algorithms for many problems in the study of Markov fields on the basis of field observations…

信息论 · 计算机科学 2018-04-04 Ulan N. Brimkulov , Chinara Jumabaeva , Kasym Baryktabasov

Bayesian inverse problems arise in various scientific and engineering domains, and solving them can be computationally demanding. This is especially the case for problems governed by partial differential equations, where the repeated…

数值分析 · 数学 2025-11-04 Juntao Yang , Jeff Adie , Simon See , Adriano Gualandi , Gianmarco Mengaldo

The stochastic heavy ball method (SHB), also known as stochastic gradient descent (SGD) with Polyak's momentum, is widely used in training neural networks. However, despite the remarkable success of such algorithm in practice, its…

机器学习 · 计算机科学 2023-02-07 Diyuan Wu , Vyacheslav Kungurtsev , Marco Mondelli

We propose a custom learning algorithm for shallow over-parameterized neural networks, i.e., networks with single hidden layer having infinite width. The infinite width of the hidden layer serves as an abstraction for the…

机器学习 · 计算机科学 2023-12-19 Alexis Teter , Iman Nodozi , Abhishek Halder

The most relevant problems in discounted reinforcement learning involve estimating the mean of a function under the stationary distribution of a Markov reward process, such as the expected return in policy evaluation, or the policy gradient…

机器学习 · 计算机科学 2023-04-17 Alberto Maria Metelli , Mirco Mutti , Marcello Restelli

In the last few decades, Markov chain Monte Carlo (MCMC) methods have been widely applied to Bayesian updating of structural dynamic models in the field of structural health monitoring. Recently, several MCMC algorithms have been developed…

应用统计 · 统计学 2026-04-29 Xianghao Meng , James L. Beck , Yong Huang , Hui Li

Markov chain Monte Carlo (MCMC) algorithms provide a very general recipe for estimating properties of complicated distributions. While their use has become commonplace and there is a large literature on MCMC theory and practice, MCMC users…

统计计算 · 统计学 2012-05-03 Murali Haran , Luke Tierney

We analyze a hybrid method that enriches coarse grid finite element solutions with fine scale fluctuations obtained from a neural network. The idea stems from the Deep Neural Network Multigrid Solver (DNN-MG), (Margenberg et al., J Comput…

数值分析 · 数学 2023-10-18 Uladzislau Kapustsin , Utku Kaya , Thomas Richter

It has become increasingly easy nowadays to collect approximate posterior samples via fast algorithms such as variational Bayes, but concerns exist about the estimation accuracy. It is tempting to build solutions that exploit approximate…

统计计算 · 统计学 2024-06-17 Leo L. Duan , Anirban Bhattacharya

Mean-field games with common noise provide a powerful framework for modeling the collective behavior of large populations subject to shared randomness, such as systemic risk in finance or environmental shocks in economics. These problems…

最优化与控制 · 数学 2025-11-13 Ruimeng Hu , Botao Jin , Mathieu Laurière , Jiacheng Zhang

For an infinite-horizon control problem, the optimal control can be represented by the stable manifold of the characteristic Hamiltonian system of Hamilton-Jacobi-Bellman (HJB) equation in a semiglobal domain. In this paper, we first…

最优化与控制 · 数学 2024-05-14 Guoyuan Chen

This paper presents a new Metropolis-adjusted Langevin algorithm (MALA) that uses convex analysis to simulate efficiently from high-dimensional densities that are log-concave, a class of probability distributions that is widely used in…

统计方法学 · 统计学 2015-04-06 Marcelo Pereyra

We introduce a novel approach to hierarchical reinforcement learning for Linearly-solvable Markov Decision Processes (LMDPs) in the infinite-horizon average-reward setting. Unlike previous work, our approach allows learning low-level and…

机器学习 · 计算机科学 2024-07-10 Guillermo Infante , Anders Jonsson , Vicenç Gómez

Although many real-world stochastic planning problems are more naturally formulated by hybrid models with both discrete and continuous variables, current state-of-the-art methods cannot adequately address these problems. We present the…

人工智能 · 计算机科学 2012-07-19 Carlos E. Guestrin , Milos Hauskrecht , Branislav Kveton