English
Related papers

Related papers: New axioms for top trading cycles

200 papers

Consider a decision-maker that can pick one out of $K$ actions to control an unknown system, for $T$ turns. The actions are interpreted as different configurations or policies. Holding the same action fixed, the system asymptotically…

Machine Learning · Computer Science 2023-02-28 Siddharth Chandak , Ilai Bistritz , Nicholas Bambos

Robust controllers ensure stability in feedback loops designed under uncertainty but at the cost of performance. Model uncertainty in time-invariant systems can be reduced by recently proposed learning-based methods, which improve the…

Systems and Control · Electrical Eng. & Systems 2023-01-18 Alexander von Rohr , Friedrich Solowjow , Sebastian Trimpe

Transient stability and critical clearing time (CCT) are important concepts in power system protection and control. This paper explores and compares various learning-based methods for predicting CCT under uncertainties arising from…

Systems and Control · Electrical Eng. & Systems 2024-09-05 Xingjian Wu , Xiaoting Wang , Xiaozhe Wang , Peter E. Caines , Jingyu Liu

The Black-Scholes-Merton model is a mathematical model for the dynamics of a financial market that includes derivative investment instruments, and its formula provides a theoretical price estimate of European-style options. The model's…

Mathematical Finance · Quantitative Finance 2023-07-04 Tongseok Lim

In Bayesian single-item auctions, a monotone bidding strategy--one that prescribes a higher bid for a higher value type--can be equivalently represented as a partition of the quantile space into consecutive intervals corresponding to…

Computer Science and Game Theory · Computer Science 2026-02-10 Junyao Zhao

In this paper, we consider the problem of periodic optimal control of nonlinear systems subject to online changing and periodically time-varying economic performance measures using model predictive control (MPC). The proposed economic MPC…

Systems and Control · Electrical Eng. & Systems 2020-10-21 Johannes Köhler , Matthias A. Müller , Frank Allgöwer

Multi-arm bandit (MAB) is a classic online learning framework that studies the sequential decision-making in an uncertain environment. The MAB framework, however, overlooks the scenario where the decision-maker cannot take actions (e.g.,…

Computer Science and Game Theory · Computer Science 2021-12-30 Zhiyuan Wang , Lin Gao , Jianwei Huang

This paper studies optimal consumption and saving decisions under uncertainty about the transition dynamics of the economic environment. We consider a general optimal savings problem in which the exogenous state governing discounting,…

Theoretical Economics · Economics 2026-03-10 Qingyin Ma , Xinxin Zhang

Bayesian optimization (BO) has become popular for sequential optimization of black-box functions. When BO is used to optimize a target function, we often have access to previous evaluations of potentially related functions. This begs the…

Machine Learning · Computer Science 2022-06-17 Zhongxiang Dai , Yizhou Chen , Haibin Yu , Bryan Kian Hsiang Low , Patrick Jaillet

We study non-rectangular robust Markov decision processes under the average-reward criterion, where the ambiguity set couples transition probabilities across states and the adversary commits to a stationary kernel for the entire horizon. We…

Optimization and Control · Mathematics 2026-03-11 Shengbo Wang , Nian Si

Given the marginal distribution information of the underlying asset price at two future times $T_1$ and $T_2$, we consider the problem of determining a model-free upper bound on the price of a class of American options that must be…

Probability · Mathematics 2023-11-03 Tongseok Lim

Model-free Reinforcement Learning (RL) works well when experience can be collected cheaply and model-based RL is effective when system dynamics can be modeled accurately. However, both assumptions can be violated in real world problems such…

Machine Learning · Computer Science 2020-05-07 Mohak Bhardwaj , Ankur Handa , Dieter Fox , Byron Boots

We initiate the study of efficient mechanism design with guaranteed good properties even when players participate in multiple different mechanisms simultaneously or sequentially. We define the class of smooth mechanisms, related to smooth…

Computer Science and Game Theory · Computer Science 2012-11-07 Vasilis Syrgkanis , Eva Tardos

We study the typical learning properties of the recently introduced Soft Margin Classifiers (SMCs), learning realizable and unrealizable tasks, with the tools of Statistical Mechanics. We derive analytically the behaviour of the learning…

Disordered Systems and Neural Networks · Physics 2009-11-07 Sebastian Risau-Gusman , Mirta B. Gordon

The MGT fluid model has been used extensively to guide designs of AQM schemes aiming to alleviate adverse effects of Internet congestion. In this paper, we provide a new analysis of a TCP/AQM system that aims to improve the accuracy of the…

Networking and Internet Architecture · Computer Science 2013-07-05 Qin Xu , Fan Li , Jinsheng Sun , Moshe Zukerman

In multi-criteria decision making (MCDM) problems, ratings are assigned to the alternatives on different criteria by the expert group. In this paper, we propose a thermodynamically consistent model for MCDM using the analogies for…

Artificial Intelligence · Computer Science 2017-03-28 Mohit Verma , J. Rajasankar

Robust reinforcement learning (RL) is to find a policy that optimizes the worst-case performance over an uncertainty set of MDPs. In this paper, we focus on model-free robust RL, where the uncertainty set is defined to be centering at a…

Machine Learning · Computer Science 2021-10-29 Yue Wang , Shaofeng Zou

This paper considers optimal control problems defined by a monotone dynamical system, a monotone cost, and monotone constraints. We identify families of such problems for which the optimal solution is bang-ride, i.e., always operates on the…

Optimization and Control · Mathematics 2023-12-15 Hamed Taghavian , Ross Drummond , Mikael Johansson

Behavioral experiments on the ultimatum game (UG) reveal that we humans prefer fair acts, which contradicts the prediction made in orthodox Economics. Existing explanations, however, are mostly attributed to exogenous factors within the…

Machine Learning · Computer Science 2026-02-04 Guozhong Zheng , Jiqiang Zhang , Xin Ou , Shengfeng Deng , Li Chen

In this paper, we provide non-averaged and transient performance guarantees for recently developed, tube-based robust economic model predictive control (MPC) schemes. In particular, we consider both tube-based MPC schemes with and without…

Systems and Control · Electrical Eng. & Systems 2022-05-17 Christian Klöppelt , Lukas Schwenkel , Frank Allgöwer , Matthias A. Müller