中文
相关论文

相关论文: Participation Incentives in Randomized Social Choi…

200 篇论文

Reallocating resources to get mutually beneficial outcomes is a fundamental problem in various multi-agent settings. While finding an arbitrary Pareto optimal allocation is generally easy, checking whether a particular allocation is Pareto…

计算机科学与博弈论 · 计算机科学 2018-05-18 Haris Aziz , Peter Biro , Jerome Lang , Julien Lesca , Jerome Monnot

Random serial dictatorship (RSD) is a randomized assignment rule that - given a set of $n$ agents with strict preferences over $n$ houses - satisfies equal treatment of equals, ex post efficiency, and strategyproofness. For $n \le 3$,…

理论经济学 · 经济学 2024-07-12 Felix Brandt , Matthias Greger , René Romen

For AI systems to be useful to humans, they must understand and act in accordance with our values and preferences. Since specifying preferences is a hard task, inverse reinforcement learning (IRL) aims to develop methods that allow for…

人工智能 · 计算机科学 2026-05-12 Karim Abdel Sadek , Mark Bedaywi , Rhys Gould , Stuart Russell

We study the classical assignment problem with initial endowments in a probabilistic framework. In this setting, each agent initially owns an object and has strict preferences over the entire set of objects, and the goal is to reassign…

理论经济学 · 经济学 2025-07-15 Sreedurga Gogulapati , Yadati Narahari , Souvik Roy , Soumyarup Sadhukhan

Election rules are formal processes that aggregate voters preferences, typically to select a single candidate, called the winner. Most of the election rules studied in the literature require the voters to rank the candidates from the most…

数据结构与算法 · 计算机科学 2019-01-31 Matthias Bentert , Piotr Skowron

A proposer requires the approval of a veto player to change a status quo. Preferences are single peaked. Proposer is uncertain about Vetoer's ideal point. We study Proposer's optimal mechanism without transfers. Vetoer is given a menu, or a…

理论经济学 · 经济学 2022-12-29 Navin Kartik , Andreas Kleiner , Richard Van Weelden

We present Self-Play Preference Optimization (SPO), an algorithm for reinforcement learning from human feedback. Our approach is minimalist in that it does not require training a reward model nor unstable adversarial training and is…

机器学习 · 计算机科学 2024-06-14 Gokul Swamy , Christoph Dann , Rahul Kidambi , Zhiwei Steven Wu , Alekh Agarwal

We consider a social choice setting with agents that are partitioned into disjoint groups, and have metric preferences over a set of alternatives. Our goal is to choose a single alternative aiming to optimize various objectives that are…

计算机科学与博弈论 · 计算机科学 2021-07-13 Elliot Anshelevich , Aris Filos-Ratsikas , Alexandros A. Voudouris

An internet network service provider manages its network with multiple objectives, such as high quality of service (QoS) and minimum computing resource usage. To achieve these objectives, a reinforcement learning-based (RL) algorithm has…

网络与互联网体系结构 · 计算机科学 2025-06-17 DongNyeong Heo , Daniela Noemi Rim , Heeyoul Choi

We study multiwinner elections with approval-based preferences. An instance of a multiwinner election consists of a set of alternatives, a population of voters---each voter approves a subset of alternatives, and the desired committee size…

计算机科学与博弈论 · 计算机科学 2019-10-15 Piotr Skowron

Using results from neurobiology on perceptual decision making and value-based decision making, the problem of decision making between lotteries is reformulated in an abstract space where uncertain prospects are mapped to corresponding…

神经元与认知 · 定量生物学 2020-01-03 Adnan Rebei

Citizens' assemblies are a form of democratic innovation in which a randomly selected panel of constituents deliberates on questions of public interest. We study a novel goal for the selection of panel members: maximizing the entropy of the…

计算机科学与博弈论 · 计算机科学 2026-04-06 Gabriel de Azevedo , Paul Gölz

It is well known that reinforcement learning can be cast as inference in an appropriate probabilistic model. However, this commonly involves introducing a distribution over agent trajectories with probabilities proportional to exponentiated…

人工智能 · 计算机科学 2021-10-07 David Tolpin , Tomer Dobkin

Incentives are more likely to elicit desired outcomes when they are designed based on accurate models of agents' strategic behavior. A growing literature, however, suggests that people do not quite behave like standard economic agents in a…

计算机科学与博弈论 · 计算机科学 2014-06-09 Arpita Ghosh , Robert Kleinberg

We propose a novel and efficient algorithm for the collaborative preference completion problem, which involves jointly estimating individualized rankings for a set of entities over a shared set of items, based on a limited number of…

机器学习 · 统计学 2016-11-16 Suriya Gunasekar , Oluwasanmi Koyejo , Joydeep Ghosh

We study fairness in social choice settings under single-peaked preferences. Construction and characterization of social choice rules in the single-peaked domain has been extensively studied in prior works. In fact, in the single-peaked…

计算机科学与博弈论 · 计算机科学 2022-07-19 Gogulapati Sreedurga , Soumyarup Sadhukhan , Souvik Roy , Yadati Narahari

Models of economic decision makers often include idealized assumptions, such as rationality, perfect foresight, and access to all relevant pieces of information. These assumptions often assure the models' internal validity, but, at the same…

综合经济学 · 经济学 2021-07-09 Patrick Reinwald , Stephan Leitner , Friederike Wall

Attention-Aware Social Choice tackles the fundamental conflict faced by some agent communities between their desire to include all members in the decision making processes and the limited time and attention that are at the disposal of the…

计算机与社会 · 计算机科学 2023-09-25 Lior Ashkenazy , Nimrod Talmon

Preference-based reinforcement learning (RL) has emerged as a new field in robot learning, where humans play a pivotal role in shaping robot behavior by expressing preferences on different sequences of state-action pairs. However,…

机器人学 · 计算机科学 2024-02-26 Simon Holk , Daniel Marta , Iolanda Leite

Learning from Preferential Feedback (LfPF) plays an essential role in training Large Language Models, as well as certain types of interactive learning agents. However, a substantial gap exists between the theory and application of LfPF…

机器学习 · 计算机科学 2024-03-29 Jonathan Colaço Carr , Prakash Panangaden , Doina Precup