中文
相关论文

相关论文: Persuasion and Incentives Through the Lens of Dual…

200 篇论文

This paper explores the potential of Lagrangian duality for learning applications that feature complex constraints. Such constraints arise in many science and engineering domains, where the task amounts to learning optimization problems…

Many algorithms in verification and automated reasoning leverage some form of duality between proofs and refutations or counterexamples. In most cases, duality is only used as an intuition that helps in understanding the algorithms and is…

编程语言 · 计算机科学 2025-01-06 Takeshi Tsukada , Hiroshi Unno , Oded Padon , Sharon Shoham

We consider Lagrangian duality based approaches to design and analyze algorithms for online energy-efficient scheduling. First, we present a primal-dual framework. Our approach makes use of the Lagrangian weak duality and convexity to…

数据结构与算法 · 计算机科学 2014-08-06 Nguyen Kim Thang

We consider large-scale Markov decision processes with an unknown cost function and address the problem of learning a policy from a finite set of expert demonstrations. We assume that the learner is not allowed to interact with the expert…

机器学习 · 计算机科学 2021-12-30 Angeliki Kamoutsi , Goran Banjac , John Lygeros

We present a unified duality approach to Bayesian persuasion. The optimal dual variable, interpreted as a price function on the state space, is shown to be a supergradient of the concave closure of the objective function at the prior…

理论经济学 · 经济学 2024-06-05 Piotr Dworczak , Anton Kolotilin

The principle of optimism in the face of uncertainty underpins many theoretically successful reinforcement learning algorithms. In this paper we provide a general framework for designing, analyzing and implementing such algorithms in the…

机器学习 · 计算机科学 2020-07-07 Gergely Neu , Ciara Pike-Burke

Despite the non-convexity of most modern machine learning parameterizations, Lagrangian duality has become a popular tool for addressing constrained learning problems. We revisit Augmented Lagrangian methods, which aim to mitigate the…

机器学习 · 计算机科学 2025-10-30 Ignacio Boero , Ignacio Hounie , Alejandro Ribeiro

Dual decomposition, and more generally Lagrangian relaxation, is a classical method for combinatorial optimization; it has recently been applied to several inference problems in natural language processing (NLP). This tutorial gives an…

计算与语言 · 计算机科学 2014-05-21 Alexander M. Rush , Michael Collins

How to optimally persuade an agent who has a private type? When elicitation is feasible, this amounts to a fairly standard principal-agent-style mechanism design problem, where the persuader employs a mechanism to first elicit the agent's…

计算机科学与博弈论 · 计算机科学 2024-11-01 Jiarui Gan , Abheek Ghosh , Nicholas Teh

We propose a modified primal-dual method for general convex optimization problems with changing constraints. We obtain properties of Lagrangian saddle points for these problems which enable us to establish convergence of the proposed…

最优化与控制 · 数学 2022-01-04 Igor Konnov

The rapid proliferation of recent Multi-Agent Systems (MAS), where Large Language Models (LLMs) and Large Reasoning Models (LRMs) usually collaborate to solve complex problems, necessitates a deep understanding of the persuasion dynamics…

人工智能 · 计算机科学 2025-09-26 Haodong Zhao , Jidong Li , Zhaomin Wu , Tianjie Ju , Zhuosheng Zhang , Bingsheng He , Gongshen Liu

Based on the complete-lattice approach, a new Lagrangian duality theory for set-valued optimization problems is presented. In contrast to previous approaches, set-valued versions for the known scalar formulas involving infimum and supremum…

最优化与控制 · 数学 2024-01-26 Andreas H. Hamel , Andreas Löhne

Designing an incentive compatible auction that maximizes expected revenue is a central problem in Auction Design. While theoretical approaches to the problem have hit some limits, a recent research direction initiated by Duetting et al.…

计算机科学与博弈论 · 计算机科学 2021-10-26 Jad Rahme , Samy Jelassi , S. Matthew Weinberg

Multi-agent systems are increasingly widespread in a range of application domains, with optimization and learning underpinning many of the tasks that arise in this context. Different approaches have been proposed to enable the cooperative…

最优化与控制 · 数学 2025-09-04 Nicola Bastianello , Luca Schenato , Ruggero Carli

In this paper we explore the role of duality principles within the problem of rotation averaging, a fundamental task in a wide range of computer vision applications. In its conventional form, rotation averaging is stated as a minimization…

计算机视觉与模式识别 · 计算机科学 2017-11-30 Anders Eriksson , Carl Olsson , Fredrik Kahl , Tat-Jun Chin

Model-free reinforcement learning methods lack an inherent mechanism to impose behavioural constraints on the trained policies. Although certain extensions exist, they remain limited to specific types of constraints, such as value…

机器学习 · 计算机科学 2025-04-28 Bram De Cooman , Johan Suykens

We study the problem of computing an optimal large language model (LLM) policy for the constrained alignment problem, where the goal is to maximize a primary reward objective while satisfying constraints on secondary utilities. Despite the…

机器学习 · 计算机科学 2025-11-27 Botong Zhang , Shuo Li , Ignacio Hounie , Osbert Bastani , Dongsheng Ding , Alejandro Ribeiro

Learning data representations that are transferable and are fair with respect to certain protected attributes is crucial to reducing unfair decisions while preserving the utility of the data. We propose an information-theoretically…

机器学习 · 计算机科学 2020-03-17 Jiaming Song , Pratyusha Kalluri , Aditya Grover , Shengjia Zhao , Stefano Ermon

Although duality is used extensively in certain fields, such as supervised learning in machine learning, it has been much less explored in others, such as reinforcement learning (RL). In this paper, we show how duality is involved in a…

机器学习 · 计算机科学 2020-07-28 Pranay Pasula

Mechanism design, a branch of economics, aims to design rules that can autonomously achieve desired outcomes in resource allocation and public decision making. The research on mechanism design using machine learning is called automated…

计算机科学与博弈论 · 计算机科学 2024-12-17 Tsuyoshi Suehara , Koh Takeuchi , Hisashi Kashima , Satoshi Oyama , Yuko Sakurai , Makoto Yokoo
‹ 上一页 1 2 3 10 下一页 ›