English
Related papers

Related papers: Incentive Non-Compatibility of Optimistic Rollups

200 papers

Reward models (RMs) play a crucial role in reinforcement learning from human feedback (RLHF), aligning model behavior with human preferences. However, existing benchmarks for reward models show a weak correlation with the performance of…

Machine Learning · Computer Science 2025-05-20 Sunghwan Kim , Dongjin Kang , Taeyoon Kwon , Hyungjoo Chae , Dongha Lee , Jinyoung Yeo

'Rich get richer' rule comforts previously often chosen actions. What is happening to the evolution of individual inclinations to choose an action when agents do interact ? Interaction tends to homogenize while each individual dynamics…

Probability · Mathematics 2020-08-05 Irene Crimaldi , Pierre-Yves Louis , Ida Germana Minelli

This paper addresses the problem of utility maximization under uncertain parameters. In contrast with the classical approach, where the parameters of the model evolve freely within a given range, we constrain them via a penalty function. We…

Optimization and Control · Mathematics 2022-03-08 Ivan Guo , Nicolas Langrené , Grégoire Loeper , Wei Ning

One obstacle to applying reinforcement learning algorithms to real-world problems is the lack of suitable reward functions. Designing such reward functions is difficult in part because the user only has an implicit understanding of the task…

Machine Learning · Computer Science 2018-11-20 Jan Leike , David Krueger , Tom Everitt , Miljan Martic , Vishal Maini , Shane Legg

We present a model predictive control (MPC) formulation to directly optimize economic criteria for linear constrained systems subject to disturbances and uncertain model parameters. The proposed formulation combines a certainty equivalent…

Systems and Control · Electrical Eng. & Systems 2024-09-11 Maximilian Degner , Raffaele Soloperto , Melanie N. Zeilinger , John Lygeros , Johannes Köhler

We introduce a simple but effective method for managing risk in model-based reinforcement learning with trajectory sampling that involves probabilistic safety constraints and balancing of optimism in the face of epistemic uncertainty and…

Machine Learning · Computer Science 2023-09-12 Marin Vlastelica , Sebastian Blaes , Cristina Pineri , Georg Martius

We consider the problem of optimal risk sharing in a pool of cooperative agents. We analyze the asymptotic behavior of the certainty equivalents and risk premia associated with the Pareto optimal risk sharing contract as the pool expands.…

Risk Management · Quantitative Finance 2017-05-01 Thomas Knispel , Roger J. A. Laeven , Gregor Svindland

Recent advances in aligning Large Language Models with human preferences have benefited from larger reward models and better preference data. However, most of these methodologies rely on the accuracy of the reward model. The reward models…

Artificial Intelligence · Computer Science 2024-11-01 Debangshu Banerjee , Aditya Gopalan

The use of equilibrium models in economics springs from the desire for parsimonious models of economic phenomena that take human reasoning into account. This approach has been the cornerstone of modern economic theory. We explain why this…

General Finance · Quantitative Finance 2008-12-02 J. Doyne Farmer , John Geanakoplos

When machine learning systems meet real world applications, accuracy is only one of several requirements. In this paper, we assay a complementary perspective originating from the increasing availability of pre-trained and regularly…

In this paper, we rigorously study the problem of cost optimisation of hybrid (mixed) institutional incentives, which are a plan of actions involving the use of reward and punishment by an external decision-maker, for maximising the level…

Populations and Evolution · Quantitative Biology 2023-10-09 M. H. Duong , C. M. Durbac , T. A. Han

Synchronization is a ubiquitous phenomenon in nonequilibrium systems. One intriguing example found in every-day life is lifts installed next to each other, that move closely and arrive almost simultaneously during a busy time. However, the…

Adaptation and Self-Organizing Systems · Physics 2026-04-02 Mitsusuke Tarama , Sakurako Tanida

Solutions of bilevel optimization problems tend to suffer from instability under changes to problem data. In the optimistic setting, we construct a lifted formulation that exhibits desirable stability properties under mild assumptions that…

Optimization and Control · Mathematics 2025-02-25 Johannes O. Royset

Reward models play a key role in aligning language model applications towards human preferences. However, this setup creates an incentive for the language model to exploit errors in the reward model to achieve high estimated reward, a…

Motivated by applications such as online labor markets we consider a variant of the stochastic multi-armed bandit problem where we have a collection of arms representing strategic agents with different performance characteristics. The…

Computer Science and Game Theory · Computer Science 2025-03-11 Seyed A. Esmaeili , Suho Shin , Aleksandrs Slivkins

In this work we investigate the inefficiency of the electricity system with strategic agents. Specifically, we prove that without a proper control the total demand of an inefficient system is at most twice the total demand of the optimal…

Computer Science and Game Theory · Computer Science 2015-09-10 Carlos Barreto , Eduardo Mojica-Nava , Nicanor Quijano

Robust optimization provides a principled framework for decision-making under uncertainty, with broad applications in finance, engineering, and operations research. In portfolio optimization, uncertainty in expected returns and covariances…

Statistical Finance · Quantitative Finance 2025-10-15 Daniel Cunha Oliveira , Grover Guzman , Nick Firoozye

Markov decision processes are widely used for planning and verification in settings that combine controllable or adversarial choices with probabilistic behaviour. The standard analysis algorithm, value iteration, only provides a lower bound…

Logic in Computer Science · Computer Science 2019-10-21 Arnd Hartmanns , Benjamin Lucien Kaminski

This paper discusses our investigation into the evolution of cooperative players in an online business environment. We explain our design of an incentive based system with its foundation over binary reputation system whose proportion of…

Computers and Society · Computer Science 2013-05-15 Sanat Kumar Bista , Keshav P Dahal , Peter I Cowling

We consider a large population of learning agents noncooperatively selecting strategies from a common set, influencing the dynamics of an exogenous system (ES) we seek to stabilize at a desired equilibrium. Our approach is to design a…

Systems and Control · Electrical Eng. & Systems 2024-09-17 Jair Certório , Nuno C. Martins , Richard J. La , Murat Arcak
‹ Prev 1 4 5 6 7 8 10 Next ›