中文
相关论文

相关论文: Agency Problems and Adversarial Bilevel Optimizati…

200 篇论文

This paper studies the stochastic optimal control of jump-diffusion processes and the associated fully nonlinear backward stochastic Hamilton--Jacobi--Bellman (BSHJB) equations. We establish the dynamic programming principle (DPP) via…

最优化与控制 · 数学 2026-05-21 Dunxiang Liang , Qingxin Meng

In this paper we provide an alternative framework to tackle the first-best Principal-Agent problem under CARA utilities. This framework leads to both a proof of existence and uniqueness of the solution to the Risk-Sharing problem under very…

风险管理 · 定量金融 2019-12-18 Jessica Martin , Anthony Réveillac

We study a bilevel optimization problem which is a zero-sum Stackelberg game. In this problem, there are two players, a leader and a follower, who pick items from a common set. Both the leader and the follower have their own…

数据结构与算法 · 计算机科学 2022-04-26 Lin Chen , Xiaoyu Wu , Guochuan Zhang

This paper studies how uncertainty about problem difficulty shapes problem-solving strategies. I develop a dynamic model where an agent solves a problem by brainstorming approaches of unknown quality and allocating a fixed effort budget…

理论经济学 · 经济学 2026-04-02 Nicholas Wu

Partially observable stochastic games provide a rich mathematical paradigm for modeling multi-agent dynamic decision making under uncertainty and partial information. However, they generally do not admit closed-form solutions and are…

最优化与控制 · 数学 2020-04-15 Yanling Chang , Chelsea C. White

A wide range of decision problems can be formulated as bilevel programs with independent followers, which as a special case include two-stage stochastic programs. These problems are notoriously difficult to solve especially when a large…

最优化与控制 · 数学 2025-09-25 Timothy C. Y. Chan , Bo Lin , Shoshanna Saxe

Consider a multi-agent systems setup in which a principal (a supervisor agent) assigns subtasks to specialized agents and aggregates their responses into a single system-level output. A core property of such systems is information…

多智能体系统 · 计算机科学 2026-02-02 Paulius Rauba , Simonas Cepenas , Mihaela van der Schaar

In this paper, further extensions of the result of the paper "A successive approximation method in functional spaces for hierarchical optimal control problems and its application to learning, arXiv:2410.20617 [math.OC], 2024" concerning a…

最优化与控制 · 数学 2024-11-26 Getachew K. Befekadu

We study a robust contract design problem with deferred inspection, in which a principal allocates a scarce resource to an agent, observes the agent's realized outcome ex post at negligible cost, and conditions transfers on this information…

理论经济学 · 经济学 2026-01-12 Halil I. Bayrak , Martin Bichler

This paper studies the problem of safe and optimal continuum deformation of a large-scale multi-agent system (MAS). We present a novel approach for MAS continuum deformation coordination that aims to achieve safe and efficient agent…

多智能体系统 · 计算机科学 2023-04-17 Harshvardhan Uppaluru , Hossein Rastgoftar

Many strategic decision-making problems, such as environment design for warehouse robots, can be naturally formulated as bi-level reinforcement learning (RL), where a leader agent optimizes its objective while a follower solves a Markov…

机器学习 · 计算机科学 2026-04-01 Mikoto Kudo , Takumi Tanabe , Akifumi Wachi , Youhei Akimoto

This paper considers a class of distributed bilevel optimization (DBO) problems with a coupled inner-level subproblem. Existing approaches typically rely on hypergradient estimations involving computationally expensive Hessian evaluation.…

最优化与控制 · 数学 2026-02-27 Youcheng Niu , Jinming Xu , Ying Sun , Li Chai , Jiming Chen

Bilevel optimization deals with nested problems in which a leader takes the first decision to minimize their objective function while accounting for a follower's best-response reaction. Constrained bilevel problems with integer variables…

最优化与控制 · 数学 2024-11-04 Justin Dumouchelle , Esther Julien , Jannis Kurtz , Elias B. Khalil

We consider a scenario in which an autonomous agent carries out a mission in a stochastic environment while passively observed by an adversary. For the agent, minimizing the information leaked to the adversary regarding its high-level…

最优化与控制 · 数学 2019-11-25 Michael Hibbard , Yagis Savas , Zhe Xu , Ufuk Topcu

We study principal-agent problems in which a principal commits to an outcome-dependent payment scheme (a.k.a. contract) so as to induce an agent to take a costly, unobservable action. We relax the assumption that the principal perfectly…

计算机科学与博弈论 · 计算机科学 2021-06-02 Matteo Castiglioni , Alberto Marchesi , Nicola Gatti

We address the problem of combined stochastic and impulse control for a market maker operating in a limit order book. The problem is formulated as a Hamilton-Jacobi-Bellman quasi-variational inequality (HJBQVI). We propose an implicit…

数理金融 · 定量金融 2025-12-25 Alexey Meteykin

We primarily consider bilevel programs where the lower level is a convex quadratic minimization problem under integer constraints. We show that it is $\Sigma_2^p$-hard to decide if the optimal objective for the leader is lesser than a given…

最优化与控制 · 数学 2024-12-23 Sriram Sankaranarayanan , V. Shubha Vatsalya

This paper tackles a multi-agent bandit setting where $M$ agents cooperate together to solve the same instance of a $K$-armed stochastic bandit problem. The agents are \textit{heterogeneous}: each agent has limited access to a local subset…

机器学习 · 计算机科学 2022-02-18 Lin Yang , Yu-zhen Janice Chen , Mohammad Hajiesmaili , John CS Lui , Don Towsley

We consider the problem of controlling the movement of multiple cooperating agents so as to minimize an uncertainty metric associated with a finite number of targets. In a one-dimensional mission space, we adopt an optimal control framework…

最优化与控制 · 数学 2016-03-15 Nan Zhou , Xi Yu , Sean B. Andersson , Christos G. Cassandras

We introduce a novel model of contracts with combinatorial actions that accounts for sequential and adaptive agent behavior. As in the standard model, a principal delegates the execution of a costly project to an agent. There are $n$…

计算机科学与博弈论 · 计算机科学 2025-04-22 Tomer Ezra , Michal Feldman , Maya Schlesinger