English
Related papers

Related papers: Efficient and Convergent Sequential Pseudo-Likelih…

200 papers

We introduce the sampling logit equilibrium (SLE), a stationary concept for population games in which agents evaluate actions using a finite sample of opponents' plays and respond according to a logit choice rule. This framework combines…

Theoretical Economics · Economics 2026-03-11 Minoru Osawa

This paper investigates the application of game-theoretic principles combined with advanced Kalman filtering techniques to enhance maritime target tracking systems. Specifically, the paper presents a two-player, imperfect information,…

Computer Science and Game Theory · Computer Science 2024-10-18 Daniel Leal , Ngoc Hung Nguyen , Alex Skvortsov , Sanjeev Arulampalam , Mahendra Piraveenan

Traditional solvable game theory and mean-field-type game theory (risk-aware games) predominantly focus on quadratic costs due to their analytical tractability. Nevertheless, they often fail to capture critical non-linearities inherent in…

Optimization and Control · Mathematics 2025-05-09 Julian Barreiro-Gomez , Tyrone E. Duncan , Bozenna Pasik-Duncan , Hamidou Tembine

The sim-to-real gap, where agents trained in a simulator face significant performance degradation during testing, is a fundamental challenge in reinforcement learning. Extansive works adopt the framework of distributionally robust RL, to…

Machine Learning · Statistics 2025-11-12 Zewu Zheng , Yuanyuan Lin

The multireference alignment problem consists of estimating a signal from multiple noisy shifted observations. Inspired by existing Unique-Games approximation algorithms, we provide a semidefinite program (SDP) based relaxation which…

Data Structures and Algorithms · Computer Science 2013-08-27 Afonso S. Bandeira , Moses Charikar , Amit Singer , Andy Zhu

Observable games are game situations that reach one of possibly many Nash equilibria. Before an instance of the game starts, an external observer does not know, a priori, what is the exact profile of actions that will occur; thus, he…

Computer Science and Game Theory · Computer Science 2022-01-04 Sandro Preto , Eduardo Fermé , Marcelo Finger

We introduce a general semiparametric clusterwise elliptical distribution to assess how latent cluster structure shapes continuous outcomes. Using a subjectwise representation, we first estimate cluster-specific mean vectors and a…

Methodology · Statistics 2026-04-10 Jen-Chieh Teng , Sheng-Hsin Fan , Chin-Tsang Chiang , Ming-Yueh Huang , Alvin Lim

We develop the fictitious play algorithm in the context of the linear programming approach for mean field games of optimal stopping and mean field games with regular control and absorption. This algorithm allows to approximate the mean…

Optimization and Control · Mathematics 2023-01-25 Roxana Dumitrescu , Marcos Leutscher , Peter Tankov

We propose a novel framework to solve risk-sensitive reinforcement learning (RL) problems where the agent optimises time-consistent dynamic spectral risk measures. Based on the notion of conditional elicitability, our methodology constructs…

Machine Learning · Computer Science 2023-05-02 Anthony Coache , Sebastian Jaimungal , Álvaro Cartea

This paper studies the last-iterate convergence properties of the exponential weights algorithm with constant learning rates. We consider a repeated interaction in discrete time, where each player uses an exponential weights algorithm…

Artificial Intelligence · Computer Science 2024-07-10 Maurizio d'Andrea , Fabien Gensbittel , Jérôme Renault

We study the problem of computing an Extensive-Form Perfect Equilibrium (EFPE) in 2-player games. This equilibrium concept refines the Nash equilibrium requiring resilience w.r.t. a specific vanishing perturbation (representing mistakes of…

Computer Science and Game Theory · Computer Science 2016-11-16 Gabriele Farina , Nicola Gatti

In this paper, we consider discrete-time partially observed mean-field games with the risk-sensitive optimality criterion. We introduce risk-sensitivity behaviour for each agent via an exponential utility function. In the game model, each…

Systems and Control · Electrical Eng. & Systems 2022-11-11 Naci Saldi , Tamer Basar , Maxim Raginsky

Although current semi-supervised medical segmentation methods can achieve decent performance, they are still affected by the uncertainty in unlabeled data and model predictions, and there is currently a lack of effective strategies that can…

Computer Vision and Pattern Recognition · Computer Science 2024-04-10 Yuanpeng He

Pattern learning in an important problem in Natural Language Processing (NLP). Some exhaustive pattern learning (EPL) methods (Bod, 1992) were proved to be flawed (Johnson, 2002), while similar algorithms (Och and Ney, 2004) showed great…

Artificial Intelligence · Computer Science 2011-04-21 Libin Shen

The paper is concerned with distributed learning in large-scale games. The well-known fictitious play (FP) algorithm is addressed, which, despite theoretical convergence results, might be impractical to implement in large-scale settings due…

Optimization and Control · Mathematics 2016-11-17 Brian Swenson , Soummya Kar , Joao Xavier

We provide a general approach to reformulating any continuous-time stochastic Stackelberg differential game under closed-loop strategies as a single-level optimisation problem with target constraints. More precisely, we consider a…

Optimization and Control · Mathematics 2026-05-14 Camilo Hernández , Nicolás Hernández Santibáñez , Emma Hubert , Dylan Possamaï

Maximum pseudo-likelihood (MPL) is a semiparametric estimation method often used to obtain the dependence parameters in copula models from data. It has been shown that despite being consistent, and in some cases efficient, MPL estimation…

Methodology · Statistics 2022-09-07 Alexandra Dias

Performative Reinforcement Learning (PRL) refers to a scenario in which the deployed policy changes the reward and transition dynamics of the underlying environment. In this work, we study multi-agent PRL by incorporating performative…

Machine Learning · Computer Science 2025-04-30 Rilind Sahitaj , Paulius Sasnauskas , Yiğit Yalın , Debmalya Mandal , Goran Radanović

This paper proposes a versatile covariate adjustment method that directly incorporates covariate balance in regression discontinuity (RD) designs. The new empirical entropy balancing method reweights the standard local polynomial RD…

Econometrics · Economics 2024-05-29 Jun Ma , Zhengfei Yu

We study the problem of computing optimal correlated equilibria (CEs) in infinite-horizon multi-player stochastic games, where correlation signals are provided over time. In this setting, optimal CEs require history-dependent policies; this…

Computer Science and Game Theory · Computer Science 2025-06-10 Jiarui Gan , Rupak Majumdar
‹ Prev 1 3 4 5 6 7 10 Next ›