English
Related papers

Related papers: New axioms for top trading cycles

200 papers

Originally motivated by default risk management applications, this paper investigates a novel problem, referred to as the profitable bandit problem here. At each step, an agent chooses a subset of the K possible actions. For each action…

Machine Learning · Statistics 2018-05-09 Mastane Achab , Stephan Clémençon , Aurélien Garivier

We consider the optimal control problem of steering an agent population to a desired distribution over an infinite horizon. This is an optimal transport problem over a dynamical system, which is challenging due to its high computational…

Optimization and Control · Mathematics 2022-03-17 Kaito Ito , Kenji Kashima

The robust constrained Markov decision process (RCMDP) is a recent task-modelling framework for reinforcement learning that incorporates behavioural constraints and that provides robustness to errors in the transition dynamics model through…

Machine Learning · Computer Science 2024-05-16 David M. Bossens

Massive machine-type communication (mMTC) is a new focus of services in fifth generation (5G) communication networks. The associated stringent delay requirement of end-to-end (E2E) service deliveries poses technical challenges. In this…

Information Theory · Computer Science 2018-07-26 Yu Gu , Qimei Cui , Qiang Ye , Weihua Zhuang

This work proposes a cooperative trading scheme for the robust optimal energy and reserve management in a multiple-microgrid (MMG) system comprising four microgrids (MGs). This scheme includes a robust optimization (RO) model which accounts…

Systems and Control · Computer Science 2018-12-04 L. P. M. I. Sampath , Ashok Krishnan , Y. S. Foo Eddy , H. B. Gooi

We model, simulate and control the guiding problem for a herd of evaders under the action of repulsive drivers. The problem is formulated in an optimal control framework, where the drivers (controls) aim to guide the evaders (states) to a…

Optimization and Control · Mathematics 2020-05-01 Dongnam Ko , Enrique Zuazua

Patch foraging involves the deliberate and planned process of determining the optimal time to depart from a resource-rich region and investigate potentially more beneficial alternatives. The Marginal Value Theorem (MVT) is frequently used…

Artificial Intelligence · Computer Science 2025-12-30 Yesid Fonseca , Manuel S. Ríos , Nicanor Quijano , Luis F. Giraldo

Traditional end-to-end contextual robust optimization models are trained for specific contextual data, requiring complete retraining whenever new contextual information arrives. This limitation hampers their use in online decision-making…

Optimization and Control · Mathematics 2025-10-20 Carlos Gamboa , Alexandre Street , Davi Valladão , Bernardo Pagnocelli

We study truthful mechanisms for approximating the Maximin-Share (MMS) allocation of agents with additive valuations for indivisible goods. Algorithmically, constant factor approximations exist for the problem for any number of agents. When…

Computer Science and Game Theory · Computer Science 2024-06-12 Ilan Reuven Cohen , Alon Eden , Talya Eden , Arsen Vasilyan

Collective migration of animals in a cohesive group is rendered possible by a strategic distribution of tasks among members: some track the travel route, which is time and energy-consuming, while the others follow the group by interacting…

Optimization and Control · Mathematics 2015-08-05 Benedetto Piccoli , Nastassia Pouradier Duteil , Benjamin Scharf

The multi-armed bandit (MAB) model is one of the most classical models to study decision-making in an uncertain environment. In this model, a player chooses one of $K$ possible arms of a bandit machine to play at each time step, where the…

Machine Learning · Computer Science 2023-06-13 Bo Li , Chi Ho Yeung

Many robotic tasks, such as human-robot interactions or the handling of fragile objects, require tight control and limitation of appearing forces and moments alongside sensible motion control to achieve safe yet high-performance operation.…

Robotics · Computer Science 2023-03-09 Janine Matschek , Johanna Bethge , Rolf Findeisen

Algorithmic trading relies on machine learning models to make trading decisions. Despite strong in-sample performance, these models often degrade when confronted with evolving real-world market regimes, which can shift dramatically due to…

Machine Learning · Computer Science 2026-01-27 Haochong Xia , Simin Li , Ruixiao Xu , Zhixia Zhang , Hongxiang Wang , Zhiqian Liu , Teng Yao Long , Molei Qin , Chuqiao Zong , Bo An

Thompson sampling (TS) is widely used in sequential decision making due to its ease of use and appealing empirical performance. However, many existing analytical and empirical results for TS rely on restrictive assumptions on reward…

Machine Learning · Computer Science 2023-06-16 Amin Karbasi , Nikki Lijing Kuang , Yi-An Ma , Siddharth Mitra

Option pricing in real markets faces fundamental challenges. The Black--Scholes--Merton (BSM) model assumes constant volatility and uses a linear generator $g(t,x,y,z)=-ry$, while lacking explicit behavioral factors, resulting in systematic…

Computational Finance · Quantitative Finance 2026-01-28 Yilun Zhang , Zheng Tang , Hexiang Sun , Yufeng Shi

In this paper we study strongly robust optimal control problems under volatility uncertainty. In the $G$-framework we adapt the stochastic maximum principle to find necessary and sufficient conditions for the existence of a strongly robust…

Optimization and Control · Mathematics 2014-04-14 Francesca Biagini , Thilo Meyer-Brandis , Bernt Øksendal , Krzysztof Paczka

Markov Decision Processes (MDPs) offer a fairly generic and powerful framework to discuss the notion of optimal policies for dynamic systems, in particular when the dynamics are stochastic. However, computing the optimal policy of an MDP…

Systems and Control · Electrical Eng. & Systems 2024-07-24 Dirk Reinhardt , Akhil S. Anand , Shambhuraj Sawant , Sebastien Gros

We address the problem of computing reliable policies in reinforcement learning problems with limited data. In particular, we compute policies that achieve good returns with high confidence when deployed. This objective, known as the…

Machine Learning · Computer Science 2021-03-01 Bahram Behzadian , Reazul Hasan Russel , Marek Petrik , Chin Pang Ho

This paper investigates the social optimum for a dynamic linear quadratic collective choice problem where a group of agents choose among multiple alternatives or destinations. The agents' common objective is to minimize the average cost of…

Optimization and Control · Mathematics 2025-06-12 Noureddine Toumi , Roland Malhamé , Jérôme Le Ny

Many physical systems have underlying safety considerations that require that the strategy deployed ensures the satisfaction of a set of constraints. Further, often we have only partial information on the state of the system. We study the…

Machine Learning · Computer Science 2022-03-30 Jiabin Lin , Xian Yeow Lee , Talukder Jubery , Shana Moothedath , Soumik Sarkar , Baskar Ganapathysubramanian