中文
相关论文

相关论文: Low-Variance and Zero-Variance Baselines for Exten…

200 篇论文

Mean-field games (MFG) have become significant tools for solving large-scale multi-agent reinforcement learning problems under symmetry. However, the assumption of exact symmetry limits the applicability of MFGs, as real-world scenarios…

计算机科学与博弈论 · 计算机科学 2024-08-28 Batuhan Yardim , Niao He

Bayesian regression games are a special class of two-player general-sum Bayesian games in which the learner is partially informed about the adversary's objective through a Bayesian prior. This formulation captures the uncertainty in regard…

机器学习 · 计算机科学 2021-10-04 Wenshuo Guo , Michael I. Jordan , Tianyi Lin

Random effects are a flexible addition to statistical models to capture structural heterogeneity in the data, such as spatial dependencies, individual differences, temporal dependencies, or non-linear effects. Testing for the presence (or…

统计方法学 · 统计学 2024-10-21 Fabio Vieira , Hongwei Zhao , Joris Mulder

Regret minimization methods are a powerful tool for learning approximate Nash equilibrium (NE) in two-player zero-sum imperfect information extensive-form games (IIEGs). We consider the problem in the interactive bandit-feedback setting…

机器学习 · 计算机科学 2023-08-21 Linjian Meng , Yang Gao

Generative Flow Networks (GFlowNets) are a family of probabilistic generative models that learn to sample compositional objects proportional to their rewards. One big challenge of GFlowNets is training them effectively when dealing with…

机器学习 · 计算机科学 2025-06-16 Zarif Ikram , Ling Pan , Dianbo Liu

Regression models are used in a wide range of applications providing a powerful scientific tool for researchers from different fields. Linear, or simple parametric, models are often not sufficient to describe complex relationships between…

机器学习 · 统计学 2021-11-24 Aliaksandr Hubin , Geir Storvik , Florian Frommlet

We present a general framework for solving a large class of learning problems with non-linear functions of classification rates. This includes problems where one wishes to optimize a non-decomposable performance metric such as the F-measure…

机器学习 · 计算机科学 2019-09-09 Harikrishna Narasimhan , Andrew Cotter , Maya Gupta

Basis Function (BF) expansions are a cornerstone of any engineer's toolbox for computational function approximation which shares connections with both neural networks and Gaussian processes. Even though BF expansions are an intuitive and…

信号处理 · 电气工程与系统科学 2024-08-15 Anton Kullberg , Frida Viset , Isaac Skog , Gustaf Hendeby

We introduce a new compositional framework for generalized variational inference, clarifying the different parts of a model, how they interact, and how they compose. We explain that both exact Bayesian inference and the loss functions…

机器学习 · 统计学 2025-03-26 Toby St Clere Smithe , Marco Perin

Counterfactual Regret Minimization (CFR) is the most popular iterative algorithm for solving zero-sum imperfect-information games. Regret-Based Pruning (RBP) is an improvement that allows poorly-performing actions to be temporarily pruned,…

计算机科学与博弈论 · 计算机科学 2016-09-13 Noam Brown , Tuomas Sandholm

We generalize gradient descent with momentum for optimization in differentiable games to have complex-valued momentum. We give theoretical motivation for our method by proving convergence on bilinear zero-sum games for simultaneous and…

机器学习 · 计算机科学 2021-06-03 Jonathan Lorraine , David Acuna , Paul Vicol , David Duvenaud

Policy Gradient methods that explore directly in parameter space are among the most effective and robust direct policy search methods and have drawn a lot of attention lately. The basic method from this field, Policy Gradients with…

机器学习 · 计算机科学 2013-12-16 Frank Sehnke

We consider zero-sum repeated games in which the players are restricted to strategies that require only a limited amount of randomness. Let $v_n$ be the max-min value of the $n$ stage game; previous works have characterized…

计算机科学与博弈论 · 计算机科学 2019-02-12 Mehrdad Valizadeh , Amin Gohari

We initiate the study of trembling-hand perfection in sequential (i.e., extensive-form) games with correlation. We introduce the extensive-form perfect correlated equilibrium (EFPCE) as a refinement of the classical extensive-form…

计算机科学与博弈论 · 计算机科学 2020-12-14 Alberto Marchesi , Nicola Gatti

A valuation for a player in a game in extensive form is an assignment of numeric values to the players moves. The valuation reflects the desirability moves. We assume a myopic player, who chooses a move with the highest valuation.…

机器学习 · 计算机科学 2007-05-23 Philippe Jehiel , Dov Samet

As neural networks are increasingly being applied to real-world applications, mechanisms to address distributional shift and sequential task learning without forgetting are critical. Methods incorporating network expansion have shown…

机器学习 · 计算机科学 2021-03-26 Vinay Kumar Verma , Kevin J Liang , Nikhil Mehta , Piyush Rai , Lawrence Carin

We study finite-memory (FM) determinacy in games on finite graphs, a central question for applications in controller synthesis, as FM strategies correspond to implementable controllers. We establish general conditions under which FM…

计算机科学与博弈论 · 计算机科学 2018-10-08 Stéphane Le Roux , Arno Pauly , Mickael Randour

We propose Fractional Policy Gradients (FPG), a reinforcement learning framework incorporating fractional calculus for long-term temporal modeling in policy optimization. Standard policy gradient approaches face limitations from Markovian…

机器学习 · 计算机科学 2025-07-02 Urvi Pawar , Kunal Telangi

Continuous-time empirical dynamic discrete choice games offer notable computational advantages over discrete-time models. This paper addresses remaining computational and econometric challenges to further improve both model solution and…

计量经济学 · 经济学 2025-11-11 Jason R. Blevins

Evolutionary game theory combines game theory and dynamical systems and is customarily adopted to describe evolutionary dynamics in multi-agent systems. In particular, it has been proven to be a successful tool to describe multi-agent…

计算机科学与博弈论 · 计算机科学 2013-04-05 Nicola Gatti , Fabio Panozzo , Marcello Restelli