English
Related papers

Related papers: Prediction with Expert Advice: a PDE Perspective

200 papers

We consider a family of learning strategies for online optimization problems that evolve in continuous time and we show that they lead to no regret. From a more traditional, discrete-time viewpoint, this continuous-time approach allows us…

Optimization and Control · Mathematics 2014-02-28 Joon Kwon , Panayotis Mertikopoulos

We consider a prediction problem with two experts and a forecaster. We assume that one of the experts is honest and makes correct prediction with probability $\mu$ at each round. The other one is malicious, who knows true outcomes at each…

Optimization and Control · Mathematics 2020-03-20 Erhan Bayraktar , H. Vincent Poor , Xin Zhang

As the dimension of a system increases, traditional methods for control and differential games rapidly become intractable, making the design of safe autonomous agents challenging in complex or team settings. Deep-learning approaches avoid…

Optimization and Control · Mathematics 2025-04-29 William Sharpless , Zeyuan Feng , Somil Bansal , Sylvia Herbert

In this paper, we present an online learning approach for two-player zero-sum linear quadratic games with unknown dynamics. We develop a framework combining regularized least squares model estimation, high probability confidence sets, and…

Systems and Control · Electrical Eng. & Systems 2026-04-06 Shanting Wang , Weihao Sun , Andreas A. Malikopoulos

The numerical solution methods for partial differential equation (PDE) solution allow obtaining a discrete field that converges towards the solution if the method is applied to the correct problem. Nevertheless, the numerical methods…

Numerical Analysis · Mathematics 2021-03-04 Alexander Hvatov

This work lies in the fusion of experimental economics and data mining. It continues author's previous work on mining behaviour rules of human subjects from experimental data, where game-theoretic predictions partially fail to work.…

Computer Science and Game Theory · Computer Science 2012-11-13 Rustam Tagiew

We study a two player repeated zero-sum game with asymmetric information introduced by Renault in which the underlying state of the game undergoes Markov evolution (parameterized by a transition probability $\frac 12\le p\le 1$). H\"orner,…

Optimization and Control · Mathematics 2017-03-16 Xavier Bressaud , Anthony Quas

This paper presents a concurrent learning-based actor-critic-identifier architecture to obtain an approximate feedback-Nash equilibrium solution to an infinite horizon N-player nonzero-sum differential game online, without requiring…

Systems and Control · Computer Science 2017-07-25 Rushikesh Kamalapurkar , Justin Klotz , Warren E. Dixon

This paper considers discounted infinite horizon mean field games by extending the probabilistic weak formulation of the game as introduced by Carmona and Lacker (2015). Under similar assumptions as in the finite horizon game, we prove…

Optimization and Control · Mathematics 2024-07-08 René Carmona , Ludovic Tangpi , Kaiwen Zhang

Two-stage bipartite matching is a fundamental problem of optimization under uncertainty introduced by Feng, Niazadeh, and Saberi (2021), who study it under the stochastic and adversarial paradigms of uncertainty. We propose a method to…

Data Structures and Algorithms · Computer Science 2024-11-06 Billy Jin , Will Ma

For zero-sum two-player continuous-time games with integral payoff and incomplete information on one side, one shows that the optimal strategy of the informed player can be computed through an auxiliary optimization problem over some…

Probability · Mathematics 2008-10-02 Pierre Cardaliaguet , Catherine Rainer

This paper develops a probabilistic numerical method for solution of partial differential equations (PDEs) and studies application of that method to PDE-constrained inverse problems. This approach enables the solution of challenging inverse…

Methodology · Statistics 2017-07-12 Jon Cockayne , Chris Oates , Tim Sullivan , Mark Girolami

Estimation of football players' skills is one of the key tasks in sports analytics. This paper introduces multiple extensions to a widely used model, expected possession value (EPV), to address some key challenges such as selection problem.…

Machine Learning · Computer Science 2024-06-04 Andrei Shelopugin

For a non-cooperative m-persons differential game, the value functions ofthe various players satisfy a system of Hamilton-Jacobi-Bellman equations.Nashequilibrium solutions in feedback form can be obtained by studying a related system of…

Optimization and Control · Mathematics 2009-01-31 Jaykov Foukzon

This paper studies a 2-players zero-sum Dynkin game arising from pricing an option on an asset whose rate of return is unknown to both players. Using filtering techniques we first reduce the problem to a zero-sum Dynkin game on a…

Probability · Mathematics 2019-05-20 Tiziano De Angelis , Fabien Gensbittel , Stéphane Villeneuve

This paper reframes approachability theory within the context of population games. Thus, whilst one player aims at driving her average payoff to a predefined set, her opponent is not malevolent but rather extracted randomly from a…

Optimization and Control · Mathematics 2014-07-16 Dario Bauso , Thomas W L Norman

Concurrent parameterized games involve a fixed yet arbitrary number of players. They are described by finite arenas in which the edges are labeled with languages that describe the possible move combinations leading from one vertex to…

Logic in Computer Science · Computer Science 2025-05-22 Nathalie Bertrand , Patricia Bouyer , Gaëtan Staquet

The paper presents numerical experiments and some theoretical developments in prediction with expert advice (PEA). One experiment deals with predicting electricity consumption depending on temperature and uses real data. As the pattern of…

Artificial Intelligence · Computer Science 2021-09-30 Vladimir V'yugin , Vladimir Trunov

We consider the Merton problem of optimizing expected power utility of terminal wealth in the case of an unobservable Markov-modulated drift. What makes the model special is that the agent is allowed to purchase costly expert opinions of…

Portfolio Management · Quantitative Finance 2024-09-19 Christoph Knochenhauer , Alexander Merkel , Yufei Zhang

Online computation is a concept to model uncertainty where not all information on a problem instance is known in advance. An online algorithm receives requests which reveal the instance piecewise and has to respond with irrevocable…

Computational Complexity · Computer Science 2023-11-28 Janosch Fuchs , Christoph Grüne , Tom Janßen