Related papers: Toxicity Bounds for Dynamic Liquidation Incentives
Reinforcement learners can attain high reward through novel unintended strategies. We study a Bayesian mitigation for general environments: we expand the agent's subjective reward range to include a large negative value $-L$, while the true…
Direct-drive is an important approach to achieving the ignition of inertial confinement fusion. To enhance implosion performance while keeping the risk of hydrodynamic instability at a low level, we have designed a procedure to optimize the…
Dynamic, risk-based pricing can systematically exclude vulnerable consumer groups from essential resources such as health insurance and consumer credit. We show that a regulator can realign private incentives with social objectives through…
This paper mathematically models a constant-function automated market maker (CFAMM) position as a portfolio of exotic options, known as perpetual American continuous-installment (CI) options. This model replicates an AMM position's delta at…
This paper studies the optimal timing to liquidate credit derivatives in a general intensity-based credit risk model under stochastic interest rate. We incorporate the potential price discrepancy between the market and investors, which is…
We propose improved fixed-design confidence bounds for the linear logistic model. Our bounds significantly improve upon the state-of-the-art bound by Li et al. (2017) via recent developments of the self-concordant analysis of the logistic…
This paper presents a novel algorithm, based on model predictive control (MPC), for the optimal guidance of a launch vehicle upper stage. The proposed strategy not only maximizes the performance of the vehicle and its robustness to external…
We assume a continuous-time price impact model similar to Almgren-Chriss but with the added assumption that the price impact parameters are stochastic processes modeled as correlated scalar Markov diffusions. In this setting, we develop…
Price-mediated contagion occurs when a positive feedback loop develops following a drop in asset prices which forces banks and other financial institutions to sell their holdings. Prior studies of such events fix the level of market…
This work considers the sample complexity of obtaining an $\varepsilon$-optimal policy in an average reward Markov Decision Process (AMDP), given access to a generative model (simulator). When the ground-truth MDP is weakly communicating,…
Experimental studies regularly show that third-party punishment (TPP) substantially exists in various settings. This study further investigates the robustness of TPP under an environment where context effects are involved. In our…
We introduce a new class of automated market maker (AMM), the \emph{partially active automated market maker} (PA-AMM). PA-AMM divides its reserves into two parts, the active and the passive parts, and uses only the active part for trading.…
We study the risk assessment of uncertain cash flows in terms of dynamic convex risk measures for processes as introduced in Cheridito, Delbaen, and Kupper (2006). These risk measures take into account not only the amounts but also the…
Automated matching engines execute millions of orders per session, yet systematic asymmetries in latency, order size, and market access compound into persistent execution disparities that erode participant trust. We formulate provably fair…
We investigate the slip boundary condition for single-phase flow past a chemically patterned surface. Molecular dynamics (MD) simulations show that modulation of fluid-solid interaction along a chemically patterned surface induces a lateral…
Automated market makers (AMMs) allocate fee revenue \textit{proportional} to the amount of liquidity investors deposit. In this paper, we study the economic consequences of the competition between passive liquidity providers (LPs) caused by…
We propose an adaptive incentive mechanism that learns the optimal incentives in environments where players continuously update their strategies. Our mechanism updates incentives based on each player's externality, defined as the difference…
This paper presents an optimal strategy for portfolio liquidation under discrete time conditions. We assume that N risky assets held will be liquidated according to the same time interval and order quantity, and the basic price processes of…
Chance constraints are widely used in stochastic model predictive control (MPC) to enforce probabilistic state and input constraints in the presence of unbounded disturbances. However, they only restrict violation probabilities and do not…
We study optimal liquidation strategies under partial information for a single asset within a finite time horizon. We propose a model tailored for high-frequency trading, capturing price formation driven solely by order flow through…