English
Related papers

Related papers: Last-iterate convergence of modified predictive me…

200 papers

This note re-visits the rolling-horizon control approach to the problem of a Markov decision process (MDP) with infinite-horizon discounted expected reward criterion. Distinguished from the classical value-iteration approach, we develop an…

Optimization and Control · Mathematics 2022-06-07 Hyeong Soo Chang

This paper presents a proper generalized decomposition (PGD) based reduced-order model of hierarchical deep-learning neural networks (HiDeNN). The proposed HiDeNN-PGD method keeps both advantages of HiDeNN and PGD methods. The automatic…

Numerical Analysis · Mathematics 2022-01-12 Lei Zhang , Ye Lu , Shaoqiang Tang , Wing Kam Liu

Over the past two decades, Machine Learning (ML) techniques have been increasingly utilized for the purpose of predicting outcomes in sport. In this paper, we provide a review of studies that have used ML for predicting results in team…

Machine Learning · Computer Science 2022-04-19 Rory Bunker , Teo Susnjak

Regret-based algorithms are highly efficient at finding approximate Nash equilibria in sequential games such as poker games. However, most regret-based algorithms, including counterfactual regret minimization (CFR) and its variants, rely on…

Machine Learning · Computer Science 2021-10-28 Chung-Wei Lee , Christian Kroer , Haipeng Luo

This paper proposes Mutation-Driven Multiplicative Weights Update (M2WU) for learning an equilibrium in two-player zero-sum normal-form games and proves that it exhibits the last-iterate convergence property in both full and noisy feedback…

Computer Science and Game Theory · Computer Science 2023-05-29 Kenshi Abe , Kaito Ariu , Mitsuki Sakamoto , Kentaro Toyoshima , Atsushi Iwasaki

In this paper, we describe a new scalable and modular material point method (MPM) code developed for solving large-scale problems in continuum mechanics. The MPM is a hybrid Eulerian-Lagrangian approach, which uses both moving material…

We present two variants of a multi-agent reinforcement learning algorithm based on evolutionary game theoretic considerations. The intentional simplicity of one variant enables us to prove results on its relationship to a system of ordinary…

Machine Learning · Computer Science 2024-05-29 Johann Bauer , Sheldon West , Eduardo Alonso , Mark Broom

This paper deals with a modified iterative projection method for approximating a solution of the hierarchical fixed point problem for a sequene of nearly nonexpansive mappings with respect to a nonexpansive mapping. It is shown that under…

Functional Analysis · Mathematics 2014-03-14 Ibrahim Karahan , Murat Ozdemir

We present a mixed-precision benchmark called HPL-MxP that uses both a lower-precision LU factorization with a non-stationary iterative refinement based on GMRES. We evaluate the numerical stability of one of the methods of generating the…

Numerical Analysis · Mathematics 2025-09-25 Jack Dongarra , Piotr Luszczek

We consider finite-horizon and infinite-horizon versions of a dynamic game with $N$ selfish players who observe their types privately and take actions that are publicly observed. Players' types evolve as conditionally independent Markov…

Optimization and Control · Mathematics 2018-03-20 Deepanshu Vasal , Abhinav Sinha , Achilleas Anastasopoulos

We propose a new policy gradient method, named homotopic policy mirror descent (HPMD), for solving discounted, infinite horizon MDPs with finite state and action spaces. HPMD performs a mirror descent type policy update with an additional…

Machine Learning · Computer Science 2022-11-30 Yan Li , Guanghui Lan , Tuo Zhao

To reduce doctors' workload, deep-learning-based automatic medical report generation has recently attracted more and more research efforts, where deep convolutional neural networks (CNNs) are employed to encode the input images, and…

Artificial Intelligence · Computer Science 2022-10-26 Wenting Xu , Zhenghua Xu , Junyang Chen , Chang Qi , Thomas Lukasiewicz

Reward models (RMs) are a critical component of reinforcement learning from human feedback (RLHF). However, conventional dense RMs are susceptible to exploitation by policy models through biases or spurious correlations, resulting in reward…

Machine Learning · Computer Science 2026-02-03 Lingling Fu , Yongfu Xue

We propose a continuous-time formulation of persistent contrastive divergence (PCD) for maximum likelihood estimation (MLE) of unnormalised densities. Our approach expresses PCD as a coupled, multiscale system of stochastic differential…

Machine Learning · Statistics 2025-10-03 Paul Felix Valsecchi Oliva , O. Deniz Akyildiz , Andrew Duncan

The use of M-estimators in generalized linear regression models in high dimensional settings requires risk minimization with hard $L_0$ constraints. Of the known methods, the class of projected gradient descent (also known as iterative hard…

Machine Learning · Computer Science 2014-10-22 Prateek Jain , Ambuj Tewari , Purushottam Kar

We introduce a novel approach to hierarchical reinforcement learning for Linearly-solvable Markov Decision Processes (LMDPs) in the infinite-horizon average-reward setting. Unlike previous work, our approach allows learning low-level and…

Machine Learning · Computer Science 2024-07-10 Guillermo Infante , Anders Jonsson , Vicenç Gómez

The extragradient method has gained popularity due to its robust convergence properties for differentiable games. Unlike single-objective optimization, game dynamics involve complex interactions reflected by the eigenvalues of the game…

Machine Learning · Computer Science 2024-02-13 Junhyung Lyle Kim , Gauthier Gidel , Anastasios Kyrillidis , Fabian Pedregosa

We deal with the numerical solution of linear partial differential equations (PDEs) with focus on the goal-oriented error estimates including algebraic errors arising by an inaccurate solution of the corresponding algebraic systems. The…

Numerical Analysis · Mathematics 2020-01-08 Vít Dolejší , Petr Tichý

Memory-Bounded Dynamic Programming (MBDP) has proved extremely effective in solving decentralized POMDPs with large horizons. We generalize the algorithm and improve its scalability by reducing the complexity with respect to the number of…

Artificial Intelligence · Computer Science 2012-06-26 Sven Seuken , Shlomo Zilberstein

As a popular and easy-to-implement machine learning method for solving differential equations, the physics-informed neural network (PINN) sometimes may fail and find poor solutions which bias against the exact ones. In this paper, we…

Analysis of PDEs · Mathematics 2023-12-22 Tao Luo , Qixuan Zhou
‹ Prev 1 8 9 10 Next ›