中文
相关论文

相关论文: Offline and Online Nonlinear Inverse Differential …

200 篇论文

We present a novel framework for {\epsilon}-optimally solving two-player zero-sum partially observable stochastic games (zs-POSGs). These games pose a major challenge due to the absence of a principled connection with dynamic programming…

计算机科学与博弈论 · 计算机科学 2025-11-17 Erwan Christian Escudie , Matthia Sabatelli , Olivier Buffet , Jilles Steeve Dibangoye

For a non-cooperative m-persons differential game, the value functions ofthe various players satisfy a system of Hamilton-Jacobi-Bellman equations.Nashequilibrium solutions in feedback form can be obtained by studying a related system of…

最优化与控制 · 数学 2009-01-31 Jaykov Foukzon

We investigate the set of Nash equilibrium payoffs for two person differential games. The main result of the paper is the characterization of the set of Nash equilibrium payoffs in the terms of nonsmooth analysis. Also we obtain the…

最优化与控制 · 数学 2015-03-17 Yurii Averboukh

Zero-sum stochastic games provide a rich model for competitive decision making. However, under general forms of state uncertainty as considered in the Partially Observable Stochastic Game (POSG), such decision making problems are still not…

人工智能 · 计算机科学 2016-06-23 Auke J. Wiggers , Frans A. Oliehoek , Diederik M. Roijers

Offline computation is an essential component in most multiscale model reduction techniques. However, there are multiscale problems in which offline procedure is insufficient to give accurate representations of solutions, due to the fact…

数值分析 · 数学 2015-04-20 Eric T. Chung , Yalchin Efendiev , Wing Tat Leung

In this paper, we investigate a class of nonzero-sum dynamic stochastic games, where players have linear dynamics and quadratic cost functions. The players are coupled in both dynamics and cost through a linear regression (weighted average)…

最优化与控制 · 数学 2020-10-20 Jalal Arabneydi , Amir G. Aghdam , Roland P. Malhamé

We introduce a network design game where the objective of the players is to design the interconnections between the nodes of two different networks $G_1$ and $G_2$ in order to maximize certain local utility functions. In this setting, each…

计算机科学与博弈论 · 计算机科学 2016-12-22 Ebrahim Moradi Shahrivar , Shreyas Sundaram

Many important real-world settings contain multiple players interacting over an unknown duration with probabilistic state transitions, and are naturally modeled as stochastic games. Prior research on algorithms for stochastic games has…

计算机科学与博弈论 · 计算机科学 2021-02-19 Sam Ganzfried

We consider two-player zero-sum differential games (ZSDGs), where the state process (dynamical system) depends on the random initial condition and the state process's distribution, and the objective functional includes the state process's…

最优化与控制 · 数学 2020-05-26 Jun Moon , Tamer Basar

We study online reinforcement learning in average-reward stochastic games (SGs). An SG models a two-player zero-sum game in a Markov environment, where state transitions and one-step payoffs are determined simultaneously by a learner and an…

机器学习 · 计算机科学 2017-12-05 Chen-Yu Wei , Yi-Te Hong , Chi-Jen Lu

The paper addresses a problem of sequential bilateral bargaining with incomplete information. We proposed a decision model that helps agents to successfully bargain by performing indirect negotiation and learning the opponent's model.…

计算机科学与博弈论 · 计算机科学 2024-09-11 Tatiana V. Guy , Jitka Homolová , Aleksej Gaj

This paper addresses a class of network games played by dynamic agents using their outputs. Unlike most existing related works, the Nash equilibrium in this work is defined by functions of agent outputs instead of full agent states, which…

最优化与控制 · 数学 2023-04-27 Meichen Guo , Claudio De Persis

In this paper we extend a popular non-cooperative network creation game (NCG) to allow for disconnected equilibrium networks. There are n players, each is a vertex in a graph, and a strategy is a subset of players to build edges to. For…

计算机科学与博弈论 · 计算机科学 2008-10-28 Ulrik Brandes , Martin Hoefer , Bobo Nick

In this paper, an aggregate game approach is proposed for the modeling and analysis of energy consumption control in smart grid. Since the electricity user's cost function depends on the aggregate load, which is unknown to the end users, an…

经济学 · 定量金融 2017-03-22 Maojiao Ye , Guoqiang Hu

Online algorithm is an important branch in algorithm design. Designing online algorithms with a bounded competitive ratio (in terms of worst-case performance) can be hard and usually relies on problem-specific assumptions. Inspired by…

机器学习 · 计算机科学 2021-11-22 Bingqian Du , Zhiyi Huang , Chuan Wu

In this paper, we consider infinite-horizon linear-quadratic cooperative differential games with output feedback information structure. We first demonstrate that, under output feedback information structure, computing Pareto optimal…

最优化与控制 · 数学 2026-05-14 Aniruddha Roy , Pavankumar Tallapragada

We present the concept of a Generalized Feedback Nash Equilibrium (GFNE) in dynamic games, extending the Feedback Nash Equilibrium concept to games in which players are subject to state and input constraints. We formalize necessary and…

最优化与控制 · 数学 2023-11-22 Forrest Laine , David Fridovich-Keil , Chih-Yuan Chiu , Claire Tomlin

This paper introduces a novel model-free and a partially model-free algorithm for inverse optimal control (IOC), also known as inverse reinforcement learning (IRL), aimed at estimating the cost function of continuous-time nonlinear…

系统与控制 · 电气工程与系统科学 2025-03-20 Hamed Jabbari Asl , Eiji Uchibe

In this paper, we present an efficient algorithm to solve online Stackelberg games, featuring multiple followers, in a follower-agnostic manner. Unlike previous works, our approach works even when leader has no knowledge about the…

最优化与控制 · 数学 2024-03-28 Chinmay Maheshwari , James Cheng , S. Shankar Sasty , Lillian Ratliff , Eric Mazumdar

This paper develops a Distributed Differentiable Dynamic Game (D3G) framework, which can efficiently solve the forward and inverse problems in multi-robot coordination. We formulate multi-robot coordination as a dynamic game, where the…

机器人学 · 计算机科学 2024-09-24 Yizhi Zhou , Wanxin Jin , Xuan Wang
‹ 上一页 1 8 9 10 下一页 ›