English
Related papers

Related papers: New axioms for top trading cycles

200 papers

This paper presents a new theory, known as robust dynamic pro- gramming, for a class of continuous-time dynamical systems. Different from traditional dynamic programming (DP) methods, this new theory serves as a fundamental tool to analyze…

Optimization and Control · Mathematics 2018-09-18 Tao Bian , Zhong-Ping Jiang

Organisms have evolved a variety of mechanisms to cope with the unpredictability of environmental conditions, and yet mainstream models of metabolic regulation are typically based on strict optimality principles that do not account for…

Molecular Networks · Quantitative Biology 2020-07-08 David S. Tourigny

Machine learning algorithms aim to find patterns from observations, which may include some noise, especially in robotics domain. To perform well even with such noise, we expect them to be able to detect outliers and discard them when…

Machine Learning · Computer Science 2020-03-04 Wendyam Eric Lionel Ilboudo , Taisuke Kobayashi , Kenji Sugimoto

In Online Learning to Rank (OLTR) the aim is to find an optimal ranking model by interacting with users. When learning from user behavior, systems must interact with users while simultaneously learning from those interactions. Unlike other…

Information Retrieval · Computer Science 2017-11-28 Harrie Oosterhuis , Maarten de Rijke

We study the problem of serving randomly arriving and delay-sensitive traffic over a multi-channel communication system with time-varying channel states and unknown statistics. This problem deviates from the classical…

Networking and Internet Architecture · Computer Science 2018-11-28 Semih Cayci , Atilla Eryilmaz

Maxmin-$\omega$ is a new threshold model, where each node in a network waits for the arrival of states from a fraction $\omega$ of neighborhood nodes before processing its own state, and subsequently transmitting it to downstream nodes.…

Physics and Society · Physics 2021-09-24 Ebrahim L. Patel

The current reinforcement learning framework focuses exclusively on performance, often at the expense of efficiency. In contrast, biological control achieves remarkable performance while also optimizing computational energy expenditure and…

Artificial Intelligence · Computer Science 2024-11-01 Devdhar Patel , Terrence Sejnowski , Hava Siegelmann

The optimal (`equilibrium') macroscopic properties of an economy with $N$ industries endowed with different technologies, $P$ commodities and one consumer are derived in the limit $N\to\infty$ with $n=N/P$ fixed using the replica method.…

Disordered Systems and Neural Networks · Physics 2008-12-02 A. De Martino , M. Marsili , I. Perez Castillo

While research of reinforcement learning applied to financial markets predominantly concentrates on finding optimal behaviours, it is worth to realize that the reinforcement learning returns $G_t$ and state value functions themselves are of…

Statistical Finance · Quantitative Finance 2024-05-21 Colin D. Grab

Who gains and who loses from a manipulable school-choice mechanism? Studying the outcomes of sincere and sophisticated students under the manipulable Boston Mechanism as compared with the strategy-proof Deferred Acceptance, we provide…

Computer Science and Game Theory · Computer Science 2020-06-12 Moshe Babaioff , Yannai A. Gonczarowski , Assaf Romm

This study investigates the development of an optimal execution strategy through reinforcement learning, aiming to determine the most effective approach for traders to buy and sell inventory within a finite time horizon. Our proposed model…

Trading and Market Microstructure · Quantitative Finance 2025-11-04 Yadh Hafsi , Edoardo Vittori

We consider the sequential decision-making problem where the mean outcome is a non-linear function of the chosen action. Compared with the linear model, two curious phenomena arise in non-linear models: first, in addition to the "learning…

Machine Learning · Statistics 2024-01-11 Nived Rajaraman , Yanjun Han , Jiantao Jiao , Kannan Ramchandran

This paper studies social optimal control of mean field LQG (linear-quadratic-Gaussian) models with uncertainty. Specially, the uncertainty is represented by a uncertain drift which is common for all agents. A robust optimization approach…

Optimization and Control · Mathematics 2019-08-06 Bing-Chang Wang , Jianhui Huang , Ji-Feng Zhang

We consider one buyer and one seller. For a bundle $(t,q)\in [0,\infty[\times [0,1]=\mathbb{Z}$, $q$ either refers to the wining probability of an object or a share of a good, and $t$ denotes the payment that the buyer makes. We define…

Computer Science and Game Theory · Computer Science 2024-10-25 Mridu Prabal Goswami

We develop a hyperparameter optimisation algorithm, Automated Budget Constrained Training (AutoBCT), which balances the quality of a model with the computational cost required to tune it. The relationship between hyperparameters, model…

Machine Learning · Statistics 2024-02-06 Lukas Cironis , Jan Palczewski , Georgios Aivaliotis

In this paper we present a framework for risk-sensitive model predictive control (MPC) of linear systems affected by stochastic multiplicative uncertainty. Our key innovation is to consider a time-consistent, dynamic risk evaluation of the…

Optimization and Control · Mathematics 2018-04-26 Sumeet Singh , Yin-Lam Chow , Anirudha Majumdar , Marco Pavone

In recent years, the dominance of machine learning in stock market forecasting has been evident. While these models have shown decreasing prediction errors, their robustness across different datasets has been a concern. A successful stock…

Computational Finance · Quantitative Finance 2025-02-18 Peiwan Wang , Chenhao Cui , Yong Li

A finite horizon optimal tracking problem is considered for linear dynamical systems subject to parametric uncertainties in the state-space matrices and exogenous disturbances. A suboptimal solution is proposed using a model predictive…

Optimization and Control · Mathematics 2022-02-08 Anilkumar Parsi , Andrea Iannelli , Roy S. Smith

We study conditions for the existence of stable and group-strategy-proof mechanisms in a many-to-one matching model with contracts if students' preferences are monotone in contract terms. We show that "equivalence", properly defined, to a…

Theoretical Economics · Economics 2021-07-13 Jan Christoph Schlegel

We investigate activities that have different periods of duration. We define the profit intensity as a measure of this economic category. The profit intensity in a repeated trading has a unique property of attaining its maximum at a fixed…

Trading and Market Microstructure · Quantitative Finance 2009-11-13 Edward W. Piotrowski , Jan Sladkowski