中文
相关论文

相关论文: An Improved Two-Party Negotiation Over Continues I…

200 篇论文

It is shown in recent studies that in a Stackelberg game the follower can manipulate the leader by deviating from their true best-response behavior. Such manipulations are computationally tractable and can be highly beneficial for the…

计算机科学与博弈论 · 计算机科学 2023-02-28 Yurong Chen , Xiaotie Deng , Jiarui Gan , Yuhao Li

The optimization of mixed-variable problems remains a significant challenge. We propose an extension of the policy-based optimization method that handles mixed-variables problems in a natural way, through a simple policy combination. This…

最优化与控制 · 数学 2025-06-17 Jonathan Viquerat

We study the selection of agents based on mutual nominations, a theoretical problem with many applications from committee selection to AI alignment. As agents both select and are selected, they may be incentivized to misrepresent their true…

计算机科学与博弈论 · 计算机科学 2025-10-23 Javier Cembrano , Felix Fischer , Max Klimm

This letter is devoted to the concept of ``instant'' model predictive control (iMPC) for linear systems. An optimization problem is formulated to express the finite-time constrained optimal regulation control, like conventional MPC. Then,…

系统与控制 · 计算机科学 2020-03-11 Keisuke Yoshida , Masaki Inoue , Takeshi Hatanaka

In this paper a novel hybrid approach for compensating the distortion of any interpolation has been proposed. In this hybrid method, a modular approach was incorporated in an iterative fashion. By using this approach we can get drastic…

多媒体 · 计算机科学 2010-09-21 A. ParandehGheibi , M. A. Akhaee , A. Ayremlou , M. A. Rahimian , F. Marvasti

Constrained decision-making is essential for designing safe policies in real-world control systems, yet simulated environments often fail to capture real-world adversities. We consider the problem of learning a policy that will maximize the…

机器学习 · 计算机科学 2026-02-10 Sourav Ganguly , Kishan Panaganti , Arnob Ghosh , Adam Wierman

Interior point methods (IPMs) are a common approach for solving linear programs (LPs) with strong theoretical guarantees and solid empirical performance. The time complexity of these methods is dominated by the cost of solving a linear…

最优化与控制 · 数学 2022-02-04 Gregory Dexter , Agniva Chowdhury , Haim Avron , Petros Drineas

Although Large Language Models (LLMs) have made remarkable progress, current preference optimization methods still struggle to align directional consistency while preserving reasoning diversity. To address this limitation, we propose…

计算与语言 · 计算机科学 2026-05-12 Mengyi Deng , Zhiwei Li , Xin Li , Tingyu Zhu , Yulan Yuan , Zhijiang Guo , Wei Wang

This paper presents SIMPOL (Simplified Policy Iteration), a modular numerical framework for solving continuous-time heterogeneous agent models. The core economic problem, the optimization of consumption and savings under idiosyncratic…

计算金融 · 定量金融 2025-09-30 Ricardo Alonzo Fernández Salguero

The proximal policy optimization (PPO) algorithm stands as one of the most prosperous methods in the field of reinforcement learning (RL). Despite its success, the theoretical understanding of PPO remains deficient. Specifically, it is…

机器学习 · 计算机科学 2023-06-09 Han Zhong , Tong Zhang

We propose a new method for the problem of controlling linear dynamical systems under partial observation and adversarial disturbances. Our new algorithm, Double Spectral Control (DSC), matches the best known regret guarantees while…

机器学习 · 计算机科学 2025-05-28 Anand Brahmbhatt , Gon Buzaglo , Sofiia Druchyna , Elad Hazan

In this paper, we propose a novel distributed algorithm for consensus optimization over networks and a robust extension tailored to deal with asynchronous agents and packet losses. Indeed, to robustly achieve dynamic consensus on the…

最优化与控制 · 数学 2025-09-04 Guido Carnevale , Nicola Bastianello , Giuseppe Notarstefano , Ruggero Carli

We discuss a dynamical systems perspective on discrete optimization. Departing from the fact that many combinatorial optimization problems can be reformulated as finding low energy spin configurations in corresponding Ising models, we…

最优化与控制 · 数学 2023-05-16 Tong Guanchun , Michael Muehlebach

An informed Advisor and an uninformed Decision-Maker, with conflicting interests, engage in repeated cheap talk communication in always new decision problems. While the Decision-Maker's optimal payoff is attainable in some subgame-perfect…

理论经济学 · 经济学 2025-08-04 Steven Kivinen , Christoph Kuzmics

Relational Markov Decision Processes are a useful abstraction for complex reinforcement learning problems and stochastic planning problems. Recent work developed representation schemes and algorithms for planning in such problems using the…

人工智能 · 计算机科学 2012-06-26 Chenggang Wang , Roni Khardon

In this note, we propose a symplectic algorithm for the stable manifolds of the Hamilton-Jacobi equations combined with an iterative procedure in [Sakamoto-van~der Schaft, IEEE Transactions on Automatic Control, 2008]. Our algorithm…

最优化与控制 · 数学 2021-08-16 Guoyuan Chen , Gaosheng Zhu

This paper proposes a novel method for designing finite-horizon discrete-valued switching signals in linear switched systems based on discreteness-promoting regularization. The inherent combinatorial optimization problem is reformulated as…

最优化与控制 · 数学 2025-05-06 Masaaki Nagahara , Takuya Ikeda , Ritsuki Hoshimoto

Alternating direction method of multiplier (ADMM) is a popular method used to design distributed versions of a machine learning algorithm, whereby local computations are performed on local data with the output exchanged among neighbors in…

机器学习 · 计算机科学 2018-06-07 Xueru Zhang , Mohammad Mahdi Khalili , Mingyan Liu

Bilateral negotiation is a complex, context-sensitive task in which human negotiators dynamically adjust anchors, pacing, and flexibility to exploit power asymmetries and informal cues. We introduce a unified mathematical framework for…

计算与语言 · 计算机科学 2025-12-16 Cheril Shah , Akshit Agarwal , Kanak Garg , Mourad Heddaya

We consider infinite horizon dynamic programming problems, where the control at each stage consists of several distinct decisions, each one made by one of several agents. In an earlier work we introduced a policy iteration algorithm, where…

最优化与控制 · 数学 2020-05-05 Dimitri Bertsekas