中文
相关论文

相关论文: Value Approximation for Two-Player General-Sum Dif…

200 篇论文

Physics informed neural networks (PINNs) have emerged as a powerful tool to provide robust and accurate approximations of solutions to partial differential equations (PDEs). However, PINNs face serious difficulties and challenges when…

机器学习 · 计算机科学 2023-07-11 Rajat Arora

With the recent surge of interest in using robotics and automation for civil purposes, providing safety and performance guarantees has become extremely important. In the past, differential games have been successfully used for the analysis…

最优化与控制 · 数学 2017-04-24 Mo Chen , Sylvia Herbert , Claire J. Tomlin

This investigation is dedicated to a two-player zero-sum stochastic differential game (SDG), where a cost function is characterized by a backward stochastic differential equation (BSDE) with a continuous and monotonic generator regarding…

最优化与控制 · 数学 2024-04-19 Guangchen Wang , Zhuangzhuang Xing

The objective of designing a control system is to steer a dynamical system with a control signal, guiding it to exhibit the desired behavior. The Hamilton-Jacobi-Bellman (HJB) partial differential equation offers a framework for optimal…

机器学习 · 计算机科学 2025-10-22 Jostein Barry-Straume , Adwait D. Verulkar , Arash Sarshar , Andrey A. Popov , Adrian Sandu

Physics-informed neural networks (PINNs) have shown promise in solving partial differential equations (PDEs) relevant to multiscale modeling, but they often fail when applied to materials with discontinuous coefficients, such as media with…

机器学习 · 计算机科学 2026-01-05 Liya Gaynutdinova , Martin Doškář , Ondřej Rokoš , Ivana Pultarová

For hyperbolic conservation laws, traditional methods and physics-informed neural networks (PINNs) often encounter difficulties in capturing sharp discontinuities and maintaining temporal consistency. To address these challenges, we…

数值分析 · 数学 2025-08-25 Yan Shen , Jingrun Chen , Keke Wu

We consider surveillance-evasion differential games, where a pursuer must try to constantly maintain visibility of a moving evader. The pursuer loses as soon as the evader becomes occluded. Optimal controls for game can be formulated as a…

人工智能 · 计算机科学 2022-03-29 Louis Ly , Yen-Hsi Richard Tsai

Dynamic programming and heuristic search are at the core of state-of-the-art solvers for sequential decision-making problems. In partially observable or collaborative settings (\eg, POMDPs and Dec-POMDPs), this requires introducing an…

计算机科学与博弈论 · 计算机科学 2022-11-16 Aurélien Delage , Olivier Buffet , Jilles Dibangoye

Learning the full family of solutions to parameterized partial differential equations (PDEs) is a central challenge to our ability to model the behavior of heterogeneous systems, with a variety of fundamental and application-oriented…

计算物理 · 物理学 2026-01-26 Milad Panahi , Giovanni Michele Porta , Monica Riva , Alberto Guadagnini

We propose physics-informed holomorphic neural networks (PIHNNs) as a method to solve boundary value problems where the solution can be represented via holomorphic functions. Specifically, we consider the case of plane linear elasticity…

计算工程、金融与科学 · 计算机科学 2024-09-30 Matteo Calafà , Emil Hovad , Allan P. Engsig-Karup , Tito Andriollo

Hamilton-Jacobi reachability (HJR) provides a value function that encodes the set of states from which a system with bounded control inputs can reach or avoid a target despite any bounded disturbance, and the corresponding robust, optimal…

系统与控制 · 电气工程与系统科学 2025-06-23 Will Sharpless , Yat Tin Chow , Sylvia Herbert

H{\infty} control of nonlinear continuous-time system depends on the solution of the Hamilton-Jacobi-Isaacs (HJI) equation, which has been proved impossible to obtain a closed-form solution due to the nonlinearity of HJI equation. In order…

系统与控制 · 电气工程与系统科学 2024-03-20 Qi Wang

This paper focuses on zero-sum stochastic differential games in the framework of forward-backward stochastic differential equations on a finite time horizon with both players adopting impulse controls. By means of BSDE methods, in…

最优化与控制 · 数学 2021-04-08 Liangquan Zhang

In this paper we propose a numerical method to obtain an approximation of Nash equilibria for multi-player non-cooperative games with a special structure. We consider the infinite horizon problem in a case which leads to a system of…

数值分析 · 数学 2016-02-19 Simone Cacace , Emiliano Cristiani , Maurizio Falcone

In various engineering and applied science applications, repetitive numerical simulations of partial differential equations (PDEs) for varying input parameters are often required (e.g., aircraft shape optimization over many design…

机器学习 · 计算机科学 2023-10-17 Woojin Cho , Kookjin Lee , Donsub Rim , Noseong Park

Recent observations have been made that bridge splitting methods arising from optimization, to the Hopf and Lax formulas for Hamilton-Jacobi Equations with Hamiltonians $H(p)$. This has produced extremely fast algorithms in computing…

最优化与控制 · 数学 2018-03-06 Alex Tong Lin , Yat Tin Chow , Stanley Osher

One of the main challenges in reinforcement learning (RL) is generalisation. In typical deep RL methods this is achieved by approximating the optimal value function with a low-dimensional representation using a deep network. While this…

机器学习 · 计算机科学 2017-11-29 Harm van Seijen , Mehdi Fatemi , Joshua Romoff , Romain Laroche , Tavian Barnes , Jeffrey Tsang

Physics-informed Neural Network (PINN) is a promising tool that has been applied in a variety of physical phenomena described by partial differential equations (PDE). However, it has been observed that PINNs are difficult to train in…

流体动力学 · 物理学 2023-07-19 E. J. R. Coutinho , M. Dall'Aqua , L. McClenny , M. Zhong , U. Braga-Neto , E. Gildin

We propose a mesh-free policy iteration framework based on physics-informed neural networks (PINNs) for solving entropy-regularized stochastic control problems. The method iteratively alternates between soft policy evaluation and…

数值分析 · 数学 2025-11-18 Yeongjong Kim , Namkyeong Cho , Minseok Kim , Yeoneung Kim

Optimal and safety-critical control are fundamental problems for stochastic systems, and are widely considered in real-world scenarios such as robotic manipulation and autonomous driving. In this paper, we consider the problem of…

系统与控制 · 电气工程与系统科学 2024-05-10 Zhuoyuan Wang , Reece Keller , Xiyu Deng , Kenta Hoshino , Takashi Tanaka , Yorie Nakahira