English
Related papers

Related papers: Non-Stationary Dynamic Pricing Via Actor-Critic In…

200 papers

The main goal of this paper is to investigate continuous-time distributed dynamic programming (DP) algorithms for networked multi-agent Markov decision problems (MAMDPs). In our study, we adopt a distributed multi-agent framework where…

Systems and Control · Electrical Eng. & Systems 2024-06-14 Donghwan Lee , Han-Dong Lim , Do Wan Kim

We study overpricing in a repeated game between two representative agents: a market maker, who controls market liquidity, and a market taker, who chooses trade quantities. Market prices evolve through the endogenous price impact of trades…

Trading and Market Microstructure · Quantitative Finance 2026-05-12 Luigi Foscari , Emanuele Guidotti , Nicolò Cesa-Bianchi , Tatjana Chavdarova , Alfio Ferrara

In this paper, a hierarchical one-leader-multi-followers game for a class of continuous-time nonlinear systems with disturbance is investigated by a novel policy iteration reinforcement learning technique in which, the game model consists…

Systems and Control · Electrical Eng. & Systems 2019-07-29 Mohammad reza Satouri , Hamed Kebriaei , Abolhassan Razminia , Mohammad javad Yazdanpanah

A novel high-frequency market-making approach in discrete time is proposed that admits closed-form solutions. By taking advantage of demand functions that are linear in the quoted bid and ask spreads with random coefficients, we model the…

Trading and Market Microstructure · Quantitative Finance 2024-05-21 Jonathan Chávez-Casillas , José E. Figueroa-López , Chuyi Yu , Yi Zhang

Two-sided online matching platforms are employed in various markets. However, agents' preferences in the current market are usually implicit and unknown, thus needing to be learned from data. With the growing availability of dynamic side…

Machine Learning · Computer Science 2024-05-30 Yuantong Li , Chi-hua Wang , Guang Cheng , Will Wei Sun

The staggering feats of AI systems have brought to attention the topic of AI Alignment: aligning a "superintelligent" AI agent's actions with humanity's interests. Many existing frameworks/algorithms in alignment study the problem on a…

Machine Learning · Computer Science 2024-10-22 Hong Jun Jeon , Benjamin Van Roy

Active learning aims to reduce annotation cost by predicting which samples are useful for a human teacher to label. However it has become clear there is no best active learning algorithm. Inspired by various philosophies about what…

Machine Learning · Computer Science 2018-10-19 Kunkun Pang , Mingzhi Dong , Yang Wu , Timothy M. Hospedales

In the Learning to Price setting, a seller posts prices over time with the goal of maximizing revenue while learning the buyer's valuation. This problem is very well understood when values are stationary (fixed or iid). Here we study the…

Computer Science and Game Theory · Computer Science 2021-06-10 Renato Paes Leme , Balasubramanian Sivan , Yifeng Teng , Pratik Worah

Efficient large-scale network allocation requires data-driven pricing mechanisms that internalize the stochastic and non-linear dynamics of user behavior. We move beyond the classic fully strategic agents to study oblivious users (agents…

Numerical Analysis · Mathematics 2026-05-28 Yixuan Li , Andersen Ang , Sebastian Stein

We describe an approximate dynamic programming (ADP) approach to compute approximations of the optimal strategies and of the minimal losses that can be guaranteed in discounted repeated games with vector-valued losses. Such games…

Computer Science and Game Theory · Computer Science 2020-10-27 Vijay Kamble , Patrick Loiseau , Jean Walrand

According to the main international reports, more pervasive industrial and business-process automation, thanks to machine learning and advanced analytic tools, will unlock more than 14 trillion USD worldwide annually by 2030. In the…

Machine Learning · Computer Science 2022-11-18 Marco Mussi , Gianmarco Genalti , Alessandro Nuara , Francesco Trovò , Marcello Restelli , Nicola Gatti

Since exchange economy considerably varies in the market assets, asset prices have become an attractive research area for investigating and modeling ambiguous and uncertain information in today markets. This paper proposes a new generative…

General Finance · Quantitative Finance 2018-03-28 Farouq Abdulaziz Masoudy

This paper provides new stability results for Action-Dependent Heuristic Dynamic Programming (ADHDP), using a control algorithm that iteratively improves an internal model of the external world in the autonomous system based on its…

Neural and Evolutionary Computing · Computer Science 2015-07-29 Yury Sokolov , Robert Kozma , Ludmilla D. Werbos , Paul J. Werbos

We consider the use of pricing as a regulatory mechanism when an unknown number of autonomous agents compete for access to a shared resource (possibly limited in volume or capacity). In standard dynamic pricing control systems, an…

Optimization and Control · Mathematics 2025-11-24 Christopher King , Homayoun Hamedmoghadam , Christos G. Cassandras , Fabian R. Wirth , Robert N. Shorten

In this paper, we develop a new method for finding an optimal biddingstrategy in sequential auctions, using a dynamic programming technique. Theexisting method assumes that the utility of a user is represented in anadditive form. Thus, the…

Computer Science and Game Theory · Computer Science 2013-01-14 Hiromitsu Hattori , Makoto Yokoo , Yuko Sakurai , Toramatsu Shintani

This work tackles the problem of robust zero-shot planning in non-stationary stochastic environments. We study Markov Decision Processes (MDPs) evolving over time and consider Model-Based Reinforcement Learning algorithms in this setting.…

Machine Learning · Computer Science 2020-01-16 Erwan Lecarpentier , Emmanuel Rachelson

Approximate Dynamic Programming (ADP) is a methodology to solve multi-stage stochastic optimization problems in multi-dimensional discrete or continuous spaces. ADP approximates the optimal value function by adaptively sampling both action…

Optimization and Control · Mathematics 2021-07-02 Vijay Kumar , Mort Webster

Pricing decisions stand out as one of the most critical tasks a company faces, particularly in today's digital economy. As with other business decision-making problems, pricing unfolds in a highly competitive and uncertain environment.…

Computer Science and Game Theory · Computer Science 2024-09-04 Daniel García Rasines , Roi Naveiro , David Ríos Insua , Simón Rodríguez Santana

A first attempt at obtaining market--directional information from a non--stationary solution of the dynamic equation "future price tends to the value that maximizes the number of shares traded per unit time" [1] is presented. We demonstrate…

Trading and Market Microstructure · Quantitative Finance 2019-03-29 Vladislav Gennadievich Malyshkin

Non-stationarity appears in many online applications such as web search and advertising. In this paper, we study the online learning to rank problem in a non-stationary environment where user preferences change abruptly at an unknown moment…

Machine Learning · Computer Science 2019-11-25 Chang Li , Maarten de Rijke