English
Related papers

Related papers: Policy Aggregation

200 papers

We argue that many general evaluation problems can be viewed through the lens of voting theory. Each task is interpreted as a separate voter, which requires only ordinal rankings or pairwise comparisons of agents to produce an overall…

Artificial Intelligence · Computer Science 2025-07-01 Marc Lanctot , Kate Larson , Yoram Bachrach , Luke Marris , Zun Li , Avishkar Bhoopchand , Thomas Anthony , Brian Tanner , Anna Koop

How AI models should deal with political topics has been discussed, but it remains challenging and requires better governance. This paper examines the governance of large language models through individual and collective deliberation,…

Computer Vision and Pattern Recognition · Computer Science 2025-07-23 Tanusree Sharma , Yujin Potter , Zachary Kilhoffer , Yun Huang , Dawn Song , Yang Wang

Even though it is unrealistic to expect citizens to pinpoint the policy implementation that they prefer from the set of alternatives, it is still possible to infer such information through an exercise of ranking the importance of policy…

Computers and Society · Computer Science 2015-09-28 Konstantinos Tserpes

We propose a new aggregation framework for approximate dynamic programming, which provides a connection with rollout algorithms, approximate policy iteration, and other single and multistep lookahead methods. The central novel…

Machine Learning · Computer Science 2019-10-08 Dimitri Bertsekas

A large number of optimization algorithms have been developed by researchers to solve a variety of complex problems in operations management area. We present a novel optimization algorithm belonging to the class of swarm intelligence…

Adaptation and Self-Organizing Systems · Physics 2016-08-05 Ilario De Vincenzo , Ilaria Giannoccaro , Giuseppe Carbone

We study mathematical models of the collaborative solving of a two-choice discrimination task. We estimate the difference between the shared performance for a group of n observers over a single person performance. Our paper is a theoretical…

Physics and Society · Physics 2013-03-13 Piotr Migdał , Michał Denkiewicz , Joanna Rcaczaszek-Leonardi , Dariusz Plewczynski

In planning problems, it is often challenging to fully model the desired specifications. In particular, in human-robot interaction, such difficulty may arise due to human's preferences that are either private or complex to model.…

Robotics · Computer Science 2021-01-01 Mahsa Ghasemi , Evan Scope Crafts , Bo Zhao , Ufuk Topcu

In a crowd forecasting system, aggregation is an algorithm that returns aggregated probabilities for each question based on the probabilities provided per question by each individual in the crowd. Various aggregation methods have been…

Applications · Statistics 2022-03-18 Yuzhong Huang , Andres Abeliuk , Fred Morstatter , Pavel Atanasov , Aram Galstyan

We consider mean-field control problems in discrete time with discounted reward, infinite time horizon and compact state and action space. The existence of optimal policies is shown and the limiting mean-field problem is derived when the…

Optimization and Control · Mathematics 2025-10-16 Nicole Bäuerle

Determining the most appropriate features for machine learning predictive models is challenging regarding performance and feature acquisition costs. In particular, global feature choice is limited given that some features will only benefit…

Machine Learning · Computer Science 2026-03-17 Gabriel Bernardino , Anders Jonsson , Patrick Clarysse , Nicolas Duchateau

Whether a population of decision-making individuals will reach a state of satisfactory decisions is a fundamental problem in studying collective behaviors. In the framework of evolutionary game theory and by means of potential functions,…

Multiagent Systems · Computer Science 2022-01-13 Negar Sakhaei , Zeinab Maleki , Pouria Ramazi

When a policy prioritizes one person over another, is it because they benefit more, or because they are preferred? This paper develops a method to uncover the values consistent with observed allocation decisions. We use machine learning…

General Economics · Economics 2022-06-03 Daniel Björkegren , Joshua E. Blumenstock , Samsun Knight

Several rules for social choice are examined from a unifying point of view that looks at them as procedures for revising a system of degrees of belief in accordance with certain specified logical constraints. Belief is here a social…

Artificial Intelligence · Computer Science 2015-05-06 Rosa Camps , Xavier Mora , Laia Saumell

Continuous-time Markov decision processes are an important class of models in a wide range of applications, ranging from cyber-physical systems to synthetic biology. A central problem is how to devise a policy to control the system in order…

Systems and Control · Computer Science 2016-06-01 Ezio Bartocci , Luca Bortolussi , Tomǎš Brázdil , Dimitrios Milios , Guido Sanguinetti

Policy optimization algorithms are crucial in many fields but challenging to grasp and implement, often due to complex calculations related to Markov decision processes and varying use of discount and average reward setups. This paper…

Systems and Control · Electrical Eng. & Systems 2025-04-07 Shuang Wu

Policy optimization on high-dimensional continuous control tasks exhibits its difficulty caused by the large variance of the policy gradient estimators. We present the action subspace dependent gradient (ASDG) estimator which incorporates…

Machine Learning · Computer Science 2019-05-30 Jiajin Li , Baoxiang Wang

Independent from the still ongoing research in measuring individual intelligence, we anticipate and provide a framework for measuring collective intelligence. Collective intelligence refers to the idea that several individuals can…

Artificial Intelligence · Computer Science 2013-07-01 Michel Halmes

The main goal of this paper is to apply the so-called policy iteration algorithm (PIA) for the long run average continuous control problem of piecewise deterministic Markov processes (PDMP's) taking values in a general Borel space and with…

Probability · Mathematics 2009-02-17 O. L. V. Costa , F. Dufour

Agentic AI aims to create systems that set their own goals, adapt proactively to change, and refine behavior through continuous experience. Recent advances suggest that, when facing multiple and unforeseen tasks, agents could benefit from…

Combinatorial preference aggregation has many applications in AI. Given the exponential nature of these preferences, compact representations are needed and ($m$)CP-nets are among the most studied ones. Sequential and global voting are two…

Artificial Intelligence · Computer Science 2019-03-28 Thomas Lukasiewicz , Enrico Malizia