中文
相关论文

相关论文: Zero-shot adaptation to order book dynamics

200 篇论文

The current paper addresses the distributed guaranteed-performance consensus design problems for general high-order linear multiagent systems with leaderless and leader-follower structures, respectively. The information about the Laplacian…

多智能体系统 · 计算机科学 2018-06-27 Jianxiang Xi , Jie Yang , Hao Liu , Tang Zheng

This paper extends the optimal-trading framework developed in arXiv:2409.03586v1 to compute optimal strategies with real-world constraints. The aim of the current paper, as with the previous, is to study trading in the context of…

交易与市场微观结构 · 定量金融 2024-09-26 Neil A. Chriss

In the setting of multi-armed trials, adaptive designs are a popular way to increase estimation efficiency, identify optimal treatments, or maximize rewards to individuals. Recent work has considered the case of estimating the effects of K…

统计方法学 · 统计学 2026-02-10 Evan T. R. Rosenman , Kristen B. Hunter

An optimal control problem is considered for a stochastic differential equation containing a state-dependent regime switching, with a recursive cost functional. Due to the non-exponential discounting in the cost functional, the problem is…

最优化与控制 · 数学 2017-12-29 Hongwei Mei , Jiongmin Yong

Developing mechanisms that flexibly adapt dialog systems to unseen tasks and domains is a major challenge in dialog research. Neural models implicitly memorize task-specific dialog policies from the training data. We posit that this…

计算与语言 · 计算机科学 2021-06-15 Shikib Mehri , Maxine Eskenazi

The discretization of overdamped Langevin dynamics, through schemes such as the Euler-Maruyama method, can be corrected by some acceptance/rejection rule, based on a Metropolis-Hastings criterion for instance. In this case, the invariant…

数值分析 · 数学 2016-07-06 Max Fathi , Gabriel Stoltz

Physics-informed neural solvers offer a promising route to model-based reinforcement learning in continuous time, where optimal feedback synthesis is governed by Hamilton--Jacobi--Bellman (HJB) equations. Practical implementations often…

机器学习 · 计算机科学 2026-05-11 Minseok Kim , Yeongjong Kim , Namkyeong Cho , Yeoneung Kim

In this study, we extend the optimal execution problem with convex market impact function studied in Kato (2014) to the case where the market impact function is S-shaped, that is, concave on $[0, \bar {x}_0]$ and convex on $[\bar {x}_0,…

数理金融 · 定量金融 2018-03-07 Takashi Kato

An architectural approach to self-adaptive systems involves runtime change of system configuration (i.e., the system's components, their bindings and operational parameters) and behaviour update (i.e., component orchestration). Thus,…

软件工程 · 计算机科学 2015-10-23 Victor Braberman , Nicolas D'Ippolito , Jeff Kramer , Daniel Sykes , Sebastian Uchitel

This paper proposes an off-policy risk-sensitive reinforcement learning based control framework for stabilization of a continuous-time nonlinear system that subjects to additive disturbances, input saturation, and state constraints. By…

系统与控制 · 电气工程与系统科学 2022-04-21 Cong Li , Qingchen Liu , Zhehua Zhou , Martin Buss , Fangzhou Liu

Demand-side management (DSM) enables distribution system operators (DSOs) to steer electricity consumption through dynamic price signals or incentive mechanisms, thereby leveraging end-users' flexibility potential for delivering grid…

最优化与控制 · 数学 2026-05-04 Silvia Cianchi , Reza Rahimi Baghbadorani , Anibal Sanjab , Sergio Grammatico

In this paper, we propose and analyse a family of generalised stochastic composite mirror descent algorithms. With adaptive step sizes, the proposed algorithms converge without requiring prior knowledge of the problem. Combined with an…

最优化与控制 · 数学 2022-11-22 Weijia Shao , Fikret Sivrikaya , Sahin Albayrak

We study market making in aggregator-routed RFQ markets where platform routing depends on slowly varying dealer performance scores. We propose a two-tier stochastic control model that separates RFQ-level price competition from a macro…

风险管理 · 定量金融 2026-03-12 Alexander Barzykin

We study mechanism design in environments where agents have private preferences and private information about a common payoff-relevant state. In such settings with multi-dimensional types, standard mechanisms fail to implement efficient…

理论经济学 · 经济学 2025-12-24 Dirk Bergemann , Marek Bojko , Paul Dütting , Renato Paes Leme , Haifeng Xu , Song Zuo

Zero-Shot learning has been shown to be an efficient strategy for domain adaptation. In this context, this paper builds on the recent work of Bucher et al. [1], which proposed an approach to solve Zero-Shot classification problems (ZSC) by…

机器学习 · 计算机科学 2016-08-29 Maxime Bucher , Stéphane Herbin , Frédéric Jurie

In this note, we study a class of indefinite stochastic McKean-Vlasov linear-quadratic (LQ in short) control problem under the control taking nonnegative values. In contrast to the conventional issue, both the classical dynamic programming…

最优化与控制 · 数学 2023-10-05 Xun Li , Liangquan Zhang

Robust control of complex engineered and biological systems hinges on the integration of feedforward and feedback mechanisms. This is exemplified in neural motor control, where feedforward muscle co-contraction complements sensory-driven…

最优化与控制 · 数学 2026-03-06 Bastien Berret , Frédéric Jean

We investigate the problem of learning an equilibrium in a generalized two-sided matching market, where agents can adaptively choose their actions based on their assigned matches. Specifically, we consider a setting in which matched agents…

机器学习 · 计算机科学 2025-06-05 Andreas Athanasopoulos , Christos Dimitrakakis

An important aspect of intelligence is the ability to adapt to a novel task without any direct experience (zero-shot), based on its relationship to previous tasks. Humans can exhibit this cognitive flexibility. By contrast, models that…

机器学习 · 计算机科学 2021-03-17 Andrew K. Lampinen , James L. McClelland

Controlling the evolution of a many-body stochastic system from a disordered reference state to a structured target ensemble, characterized empirically through samples, arises naturally in non-equilibrium statistical mechanics and…

统计力学 · 物理学 2026-04-10 Haiqian Yang , Vishaal Krishnan , Sumit Sinha , L. Mahadevan