中文
相关论文

相关论文: Approximately Gaussian Replicator Flows: Nonconvex…

200 篇论文

We consider seeking generalized Nash equilibria (GNE) for noncooperative games with coupled nonlinear constraints over networks. We first revisit a well-known gradientplay dynamics from a passivity-based perspective, and address that the…

最优化与控制 · 数学 2024-08-23 Weijian Li , Lacra Pavel

We propose computationally tractable accelerated first-order methods for Riemannian optimization, extending the Nesterov accelerated gradient (NAG) method. For both geodesically convex and geodesically strongly convex objective functions,…

最优化与控制 · 数学 2025-08-12 Jungbin Kim , Insoon Yang

In this letter, we study distributed optimization and Nash equilibrium-seeking dynamics from a contraction theoretic perspective. Our first result is a novel bound on the logarithmic norm of saddle matrices. Second, for distributed gradient…

最优化与控制 · 数学 2023-09-25 Anand Gokhale , Alexander Davydov , Francesco Bullo

We focus on the problem of \emph{Answer-Level Fine-Tuning} (ALFT), where the goal is to optimize a language model based on the correctness or properties of its final answers, rather than the specific reasoning traces used to produce them.…

机器学习 · 计算机科学 2026-05-01 Mehryar Mohri , Jon Schneider , Yifan Wu

This paper studies approximate solutions to large-scale linear quadratic stochastic games with homogeneous nodal dynamics parameters and heterogeneous network couplings within the graphon mean field game framework in [2]-[4]. A graphon…

系统与控制 · 电气工程与系统科学 2021-10-22 Shuang Gao , Peter E. Caines , Minyi Huang

Almost all of the work in graphical models for game theory has mirrored previous work in probabilistic graphical models. Our work considers the opposite direction: Taking advantage of recent advances in equilibrium computation for…

人工智能 · 计算机科学 2017-10-10 Luis E. Ortiz , Boshen Wang , Ze Gong

Entropic optimal transport problems are regularized versions of optimal transport problems. These models play an increasingly important role in machine learning and generative modelling. For finite spaces, these problems are commonly solved…

机器学习 · 统计学 2025-12-30 O. Deniz Akyildiz , Pierre Del Moral , Joaquín Miguez

This paper studies the continuous-time dynamics generated by control-theoretic Lagrangian methods for equality-constrained optimization. In particular, we consider dynamics induced by proportional-integral and feedback linearization…

最优化与控制 · 数学 2026-05-26 Simone Pirrera , Francesco Ripa , Daniele Astolfi , Vito Cerone , Sophie M. Fosson , Diego Regruto

There has been substantial progress on finding game-theoretic equilibria. Most of that work has focused on games with finite, discrete action spaces. However, many games involving space, time, money, and other fine-grained quantities have…

计算机科学与博弈论 · 计算机科学 2025-10-28 Carlos Martin , Tuomas Sandholm

The recent mean field game (MFG) formalism facilitates otherwise intractable computation of approximate Nash equilibria in many-agent settings. In this paper, we consider discrete-time finite MFGs subject to finite-horizon objectives. We…

多智能体系统 · 计算机科学 2022-07-11 Kai Cui , Heinz Koeppl

While Online Gradient Descent and other no-regret learning procedures are known to efficiently converge to a coarse correlated equilibrium in games where each agent's utility is concave in their own strategy, this is not the case when…

计算机科学与博弈论 · 计算机科学 2025-04-22 Yang Cai , Constantinos Daskalakis , Haipeng Luo , Chen-Yu Wei , Weiqiang Zheng

We propose a nonparametric density estimator based on the Gaussian process (GP) and derive three novel closed form learning algorithms based on Fisher divergence (FD) score matching. The density estimator is formed by multiplying a base…

机器学习 · 计算机科学 2025-11-17 John Paisley , Wei Zhang , Brian Barr

Recently, invariant risk minimization (IRM) (Arjovsky et al.) was proposed as a promising solution to address out-of-distribution (OOD) generalization. In Ahuja et al., it was shown that solving for the Nash equilibria of a new class of…

机器学习 · 计算机科学 2020-10-30 Kartik Ahuja , Karthikeyan Shanmugam , Amit Dhurandhar

The properties of lattice-based structures can be enhanced by varying their geometric parameters in a graded manner, and the gradation can be tailored to extremize a particular objective. In this manuscript, we propose a non-gradient-based…

计算物理 · 物理学 2026-04-07 Piyush Agrawal , Manish Agrawal

This paper investigates posterior sampling algorithms for competitive reinforcement learning (RL) in the context of general function approximations. Focusing on zero-sum Markov games (MGs) under two critical settings, namely self-play and…

机器学习 · 计算机科学 2023-11-01 Shuang Qiu , Ziyu Dai , Han Zhong , Zhaoran Wang , Zhuoran Yang , Tong Zhang

We present an optimization algorithm that can identify a global minimum of a potentially nonconvex smooth function with high probability, assuming the Gibbs measure of the potential satisfies a logarithmic Sobolev inequality. Our…

最优化与控制 · 数学 2025-09-16 Daniel Cortild , Claire Delplancke , Nadia Oudjane , Juan Peypouquet

We present new algorithms for inverse reinforcement learning (IRL, or inverse optimal control) in convex optimization settings. We argue that finite-space IRL can be posed as a convex quadratic program under a Bayesian inference framework…

机器学习 · 计算机科学 2013-01-22 Qifeng Qiao , Peter A. Beling

In this paper, we deal with the equilibrium selection problem, which amounts to steering a population of individuals engaged in strategic game-theoretic interactions to a desired collective behavior. In the literature, this problem has been…

系统与控制 · 电气工程与系统科学 2025-11-11 Lorenzo Zino , Mengbin Ye , Giuseppe Carlo Calafiore , Alessandro Rizzo

This paper studies two fundamental problems in regularized Graphon Mean-Field Games (GMFGs). First, we establish the existence of a Nash Equilibrium (NE) of any $\lambda$-regularized GMFG (for $\lambda\geq 0$). This result relies on weaker…

计算机科学与博弈论 · 计算机科学 2023-10-13 Fengzhuo Zhang , Vincent Y. F. Tan , Zhaoran Wang , Zhuoran Yang

We review convergence and behavior of stochastic gradient descent for convex and nonconvex optimization, establishing various conditions for convergence to zero of the variance of the gradient of the objective function, and presenting a…

最优化与控制 · 数学 2025-03-06 Kevin Buck , Jessica Babyak , Paolo Piersanti , Kevin Zumbrun , Christiane Gallos , Dorothea Gallos