中文
相关论文

相关论文: Deep Learning for Population-Dependent Controls in…

200 篇论文

We present APAC-Net, an alternating population and agent control neural network for solving stochastic mean field games (MFGs). Our algorithm is geared toward high-dimensional instances of MFGs that are beyond reach with existing solution…

机器学习 · 计算机科学 2023-07-17 Alex Tong Lin , Samy Wu Fung , Wuchen Li , Levon Nurbekyan , Stanley J. Osher

We propose a single-level numerical approach to solve Stackelberg mean field game (MFG) problems. In Stackelberg MFG, an infinite population of agents play a non-cooperative game and choose their controls to optimize their individual…

最优化与控制 · 数学 2024-04-24 Gokce Dayanikli , Mathieu Lauriere

In this paper, we leverage the rapid advances in imitation learning, a topic of intense recent focus in the Reinforcement Learning (RL) literature, to develop new sample complexity results and performance guarantees for data-driven Model…

最优化与控制 · 数学 2022-10-18 Kwangjun Ahn , Zakaria Mhammedi , Horia Mania , Zhang-Wei Hong , Ali Jadbabaie

The success of deep learning ignited interest in whether the brain learns hierarchical representations using gradient-based learning. However, current biologically plausible methods for gradient-based credit assignment in deep neural…

神经与进化计算 · 计算机科学 2022-06-23 Alexander Meulemans , Matilde Tristany Farinha , Maria R. Cervera , João Sacramento , Benjamin F. Grewe

Decentralized learning and optimization is a central problem in control that encompasses several existing and emerging applications, such as federated learning. While there exists a vast literature on this topic and most methods centered…

机器学习 · 计算机科学 2023-03-21 Vishnu Pandi Chellapandi , Antesh Upadhyay , Abolfazl Hashemi , Stanislaw H /. Zak

In this paper, how to successfully and efficiently condition a target population of agents towards consensus is discussed. To overcome the curse of dimensionality, the mean field formulation of the consensus control problem is considered.…

最优化与控制 · 数学 2022-07-20 Giacomo Albi , Sara Bicego , Dante Kalise

We develop a scalable algorithm for mean field control problems with kernel interactions by combining particle system simulations with random Fourier feature approximations. The method replaces the quadratic-cost kernel evaluations by…

最优化与控制 · 数学 2026-05-25 Zhongyuan Cao , Kaustav Das , Nicolas Langrené , Mathieu Laurière

We present differentiable predictive control (DPC), a method for learning constrained neural control policies for linear systems with probabilistic performance guarantees. We employ automatic differentiation to obtain direct policy…

系统与控制 · 电气工程与系统科学 2022-01-28 Jan Drgona , Aaron Tuor , Draguna Vrabie

The goal of this work is to obtain optimal rates for the convergence problem in mean field control. Our analysis covers cases where the solutions to the limiting problem may not be unique nor stable. Equivalently the value function of the…

最优化与控制 · 数学 2023-05-16 Samuel Daudin , François Delarue , Joe Jackson

This article investigates synthetic model-predictive control (MPC) problems to demonstrate that an increased precision of the internal prediction model (PM) automatially entails an improvement of the controller as a whole. In contrast to…

机器学习 · 计算机科学 2023-08-30 L. Féret , A. Gepperth , S. Lambeck

Imitation learning considerably simplifies policy synthesis compared to alternative approaches by exploiting access to expert demonstrations. For such imitation policies, errors away from the training samples are particularly critical. Even…

机器学习 · 计算机科学 2024-03-19 Kaustubh Sridhar , Souradeep Dutta , Dinesh Jayaraman , James Weimer , Insup Lee

We establish an algorithm to learn feedback maps from data for a class of robust model predictive control (MPC) problems. The algorithm accounts for the approximation errors due to the learning directly at the synthesis stage, ensuring…

最优化与控制 · 数学 2025-10-16 Siddhartha Ganguly , Shubham Gupta , Debasish Chatterjee

We propose a neural network approach to model general interaction dynamics and an adjoint based stochastic gradient descent algorithm to calibrate its parameters. The parameter calibration problem is considered as optimal control problem…

最优化与控制 · 数学 2021-02-01 Simone Göttlich , Claudia Totzeck

In this paper we model the role of a government of a large population as a mean field optimal control problem. Such control problems are constrainted by a PDE of continuity-type, governing the dynamics of the probability distribution of the…

最优化与控制 · 数学 2016-08-08 Giacomo Albi , Young-Pil Choi , Massimo Fornasier , Dante Kalise

This paper investigates the social optimality of linear quadratic mean field control systems with unmodeled dynamics. The objective of agents is to optimize the social cost, which is the sum of costs of all agents. By variational analysis…

最优化与控制 · 数学 2020-11-30 Bing-Chang Wang , Yong Liang

In this paper, we present an approach to neural network mean-field-type control and its stochastic stability analysis by means of adversarial inputs (aka adversarial attacks). This is a class of data-driven mean-field-type control where the…

最优化与控制 · 数学 2022-10-04 Julian Barreiro-Gomez , Salah Eddine Choutri , Boualem Djehiche

Automatic analysis of highly crowded people has attracted extensive attention from computer vision research. Previous approaches for crowd counting have already achieved promising performance across various benchmarks. However, to deal with…

计算机视觉与模式识别 · 计算机科学 2020-02-18 Xiaowen Shi , Xin Li , Caili Wu , Shuchen Kong , Jing Yang , Liang He

Meta-learning usually refers to a learning algorithm that learns from other learning algorithms. The problem of uncertainty in the predictions of neural networks shows that the world is only partially predictable and a learned neural…

机器学习 · 计算机科学 2023-02-27 Yuwei Sun

In this paper we propose a Deep Learning architecture to approximate diffeomorphisms diffeotopic to the identity. We consider a control system of the form $\dot x = \sum_{i=1}^lF_i(x)u_i$, with linear dependence in the controls, and we use…

最优化与控制 · 数学 2023-11-20 Alessandro Scagliotti

Crowd counting is one of the core tasks in various surveillance applications. A practical system involves estimating accurate head counts in dynamic scenarios under different lightning, camera perspective and occlusion states. Previous…

计算机视觉与模式识别 · 计算机科学 2018-06-27 Li Wang , Weiyuan Shao , Yao Lu , Hao Ye , Jian Pu , Yingbin Zheng