中文
相关论文

相关论文: Game-Theoretic Algorithms for Conditional Moment M…

200 篇论文

We study offline multi-agent reinforcement learning (RL) in Markov games, where the goal is to learn an approximate equilibrium -- such as Nash equilibrium and (Coarse) Correlated Equilibrium -- from an offline dataset pre-collected from…

机器学习 · 计算机科学 2023-02-07 Yuheng Zhang , Yu Bai , Nan Jiang

We give sublinear-time approximation algorithms for some optimization problems arising in machine learning, such as training linear classifiers and finding minimum enclosing balls. Our algorithms can be extended to some kernelized versions…

机器学习 · 计算机科学 2010-10-22 Kenneth L. Clarkson , Elad Hazan , David P. Woodruff

In common-interest stochastic games all players receive an identical payoff. Players participating in such games must learn to coordinate with each other in order to receive the highest-possible value. A number of reinforcement learning…

人工智能 · 计算机科学 2011-06-28 R. I. Brafman , M. Tennenholtz

Often machine learning and statistical models will attempt to describe the majority of the data. However, there may be situations where only a fraction of the data can be fit well by a linear regression model. Here, we are interested in a…

机器学习 · 计算机科学 2021-11-16 Brendan Juba , Leda Liang

An improved exponential time algorithm for Energy Games and Mean Payoff Games has been recently proposed in ICALP 19. The new algorithm prevents some of the repetitive operations performed by the classic value iteration algorithm of Brim et…

数据结构与算法 · 计算机科学 2023-10-09 Peter Austin , Daniele Dell'Erba

This paper investigates the application of game-theoretic principles combined with advanced Kalman filtering techniques to enhance maritime target tracking systems. Specifically, the paper presents a two-player, imperfect information,…

计算机科学与博弈论 · 计算机科学 2024-10-18 Daniel Leal , Ngoc Hung Nguyen , Alex Skvortsov , Sanjeev Arulampalam , Mahendra Piraveenan

In the literature on game-theoretic equilibrium finding, focus has mainly been on solving a single game in isolation. In practice, however, strategic interactions -- ranging from routing problems to online advertising auctions -- evolve…

计算机科学与博弈论 · 计算机科学 2023-03-02 Keegan Harris , Ioannis Anagnostides , Gabriele Farina , Mikhail Khodak , Zhiwei Steven Wu , Tuomas Sandholm

This paper explores advanced topics in complex multi-agent systems building upon our previous work. We examine four fundamental challenges in Multi-Agent Reinforcement Learning (MARL): non-stationarity, partial observability, scalability…

多智能体系统 · 计算机科学 2024-12-31 Neil De La Fuente , Miquel Noguer i Alonso , Guim Casadellà

The Boltzmann equation, a fundamental equation in kinetic theory, serves as a bridge between microscopic particle dynamics and macroscopic continuum mechanics. However, deriving closed macroscopic moment systems from the Boltzmann equation…

数值分析 · 数学 2025-07-29 Juntao Huang , Liu Liu , Kunlun Qi , Jiayu Wan

We discuss in detail the derivation of stochastic differential equations for the continuum time limit of the Minority Game. We show that all properties of the Minority Game can be understood by a careful theoretical analysis of such…

统计力学 · 物理学 2009-11-07 M. Marsili , D. Challet

Learning in games has emerged as a powerful tool for machine learning with numerous applications. Quantum games model interactions between strategic players who have access to quantum resources, and several recent works have studied…

计算机科学与博弈论 · 计算机科学 2025-04-09 Wayne Lin , Georgios Piliouras , Ryann Sim , Antonios Varvitsiotis

We investigate discrete-time mean-variance portfolio selection problems viewed as a Markov decision process. We transform the problems into a new model with deterministic transition function for which the Bellman optimality equation holds.…

最优化与控制 · 数学 2025-09-23 Nicole Bäuerle , Anna Jaśkiewicz

Goal-models (GM) have been used in adaptive systems engineering for their ability to capture the different ways to fulfill the requirements. Contextual GM (CGM) extend these models with the notion of context and context-dependent…

软件工程 · 计算机科学 2015-03-25 Felipe Pontes Guimarães , Genaina Nunes Rodrigues , Raian Ali , Daniel Macêdo Batista

Deriving competitive, distributed solutions to multi-agent problems is crucial for many developing application domains; Game theory has emerged as a useful framework to design such algorithms. However, much of the attention within this…

系统与控制 · 电气工程与系统科学 2024-06-27 Rohit Konda , Rahul Chandan , David Grimsman , Jason R. Marden

Traditional reinforcement learning (RL) aims to maximize the expected total reward, while the risk of uncertain outcomes needs to be controlled to ensure reliable performance in a risk-averse setting. In this paper, we consider the problem…

机器学习 · 计算机科学 2023-01-18 Xian Yu , Siqian Shen

Many modern computational approaches to classical problems in quantitative finance are formulated as empirical loss minimization (ERM), allowing direct applications of classical results from statistical machine learning. These methods,…

机器学习 · 统计学 2022-09-27 A. Max Reppen , H. Mete Soner

Two algorithms are proposed, analyzed, and tested for solving continuous optimization problems with nonlinear equality constraints. Each is an extension of a stochastic momentum-based method from the unconstrained setting to the setting of…

最优化与控制 · 数学 2026-01-21 Qi Wang , Christian Piermarini , Yunlang Zhu , Frank E. Curtis

This paper deals with N-person nonzero-sum discrete-time Markov games under a probability criterion, in which the transition probabilities and reward functions are allowed to vary with time. Differing from the existing works on the expected…

概率论 · 数学 2025-05-16 Xin Guo , Xin Wen

A multi-variable adaptive controller is derived as the explicit solution to a minimax dynamic game. The minimizing player selects the control action as a function of past state measurements and inputs. The maximizing player selects…

最优化与控制 · 数学 2025-12-01 Anders Rantzer

Many decision problems in economics, information technology, and industry can be transformed to an optimal stopping of adapted random vectors with some utility function over the set of Markov times with respect to filtration build by the…

最优化与控制 · 数学 2020-11-04 Krzysztof Szajowski