English
Related papers

Related papers: Non-equilibrium time-dependent solution to discret…

200 papers

In this article we present a general framework for non-concave robust stochastic control problems under model uncertainty in a discrete time finite horizon setting. Our framework allows to consider a variety of different path-dependent…

Optimization and Control · Mathematics 2025-05-06 Ariel Neufeld , Julian Sester

Fitting models to data is an important part of the practice of science. Advances in machine learning have made it possible to fit more -- and more complex -- models, but have also exacerbated a problem: when multiple models fit the data…

Methodology · Statistics 2025-10-27 Alexandre René , André Longtin

We develop a regression based primal-dual martingale approach for solving finite time horizon MDPs with general state and action space. As a result, our method allows for the construction of tight upper and lower biased approximations of…

Numerical Analysis · Mathematics 2022-10-05 Denis Belomestny , John Schoenmakers

We consider a sequential decision making problem where the agent faces the environment characterized by the stochastic discrete events and seeks an optimal intervention policy such that its long-term reward is maximized. This problem exists…

Machine Learning · Computer Science 2022-12-29 Chao Qu , Xiaoyu Tan , Siqiao Xue , Xiaoming Shi , James Zhang , Hongyuan Mei

A solution to control for nonresponse bias consists of multiplying the design weights of respondents by the inverse of estimated response probabilities to compensate for the nonrespondents. Maximum likelihood and calibration are two…

Methodology · Statistics 2023-10-27 Caren Hasler

Current approaches to model-based offline reinforcement learning often incorporate uncertainty-based reward penalization to address the distributional shift problem. These approaches, commonly known as pessimistic value iteration, use Monte…

Machine Learning · Computer Science 2025-01-17 Abdullah Akgül , Manuel Haußmann , Melih Kandemir

We propose a data-driven method to learn the time-dependent probability density of a multivariate stochastic process from sample paths, assuming that the initial probability density is known and can be evaluated. Our method uses a novel…

Machine Learning · Statistics 2025-06-19 Agnimitra Dasgupta , Javier Murgoitio-Esandi , Ali Fardisi , Assad A Oberai

We continue our study of the linear response of a nonequilibrium system. This Part II concentrates on models of open and driven inertial dynamics but the structure and the interpretation of the result remain unchanged: the response can be…

Statistical Mechanics · Physics 2010-05-02 Marco Baiesi , Eliran Boksenbojm , Christian Maes , Bram Wynants

In a distributed algorithm, multiple processes, or agents, work toward a common goal. More often than not, the actions of some agents are dependent on the previous execution (if not also on the outcome) of the actions of other agents. The…

Multiagent Systems · Computer Science 2012-06-12 Yannai A. Gonczarowski

Suppose an online platform wants to compare a treatment and control policy, e.g., two different matching algorithms in a ridesharing system, or two different inventory management algorithms in an online retail site. Standard randomized…

Methodology · Statistics 2022-12-27 Peter Glynn , Ramesh Johari , Mohammad Rasouli

A researcher observes a finite sequence of choices made by multiple agents in a binary-state environment. Agents maximize expected utilities that depend on their chosen alternative and the unknown underlying state. Agents learn about the…

Theoretical Economics · Economics 2021-05-11 Rahul Deb , Ludovic Renou

We propose a learning-based robust predictive control algorithm that compensates for significant uncertainty in the dynamics for a class of discrete-time systems that are nominally linear with an additive nonlinear component. Such systems…

Systems and Control · Electrical Eng. & Systems 2022-12-05 Rohan Sinha , James Harrison , Spencer M. Richards , Marco Pavone

We study collective decision-making in a model of human groups, with network interactions, performing two alternative choice tasks. We focus on the speed-accuracy tradeoff, i.e., the tradeoff between a quick decision and a reliable…

Optimization and Control · Mathematics 2014-02-18 Vaibhav Srivastava , Naomi Ehrich Leonard

We develop a theory for continuous-time non-Markovian stochastic control problems which are inherently time-inconsistent. Their distinguishing feature is that the classical Bellman optimality principle no longer holds. Our formulation is…

Optimization and Control · Mathematics 2021-08-03 Camilo Hernández , Dylan Possamaï

As a schematic model of the complexity economic agents are confronted with, we introduce the ``SK-game'', a discrete time binary choice model inspired from mean-field spin-glasses. We show that even in a completely static environment,…

Statistical Mechanics · Physics 2024-08-27 Jerome Garnier-Brun , Michael Benzaquen , Jean-Philippe Bouchaud

We study a discrete-time portfolio selection problem with partial information and maxi\-mum drawdown constraint. Drift uncertainty in the multidimensional framework is modeled by a prior probability distribution. In this Bayesian framework,…

Portfolio Management · Quantitative Finance 2020-11-02 Carmine De Franco , Johann Nicolle , Huyên Pham

The time evolution of correlation functions in statistical systems is described by an exact functional differential equation for the corresponding generating functionals. This allows for a systematic discussion of non-equilibrium physics…

High Energy Physics - Theory · Physics 2009-10-30 Christof Wetterich

In this work, we develop a game-theoretic modeling of the interaction between a human operator and an autonomous decision aid when they collaborate in a multi-agent task allocation setting. In this setting, we propose a decision aid that is…

Multiagent Systems · Computer Science 2021-12-21 Larkin Heintzman , Ryan K. Williams

We extend Kirman's model by introducing variable event time scale. The proposed flexible time scale is equivalent to the variable trading activity observed in financial markets. Stochastic version of the extended Kirman's agent based model…

Statistical Finance · Quantitative Finance 2011-12-23 Aleksejus Kononovicius , Vygintas Gontis

Robust Model Predictive Control (MPC) for nonlinear systems is a problem that poses significant challenges as highlighted by the diversity of approaches proposed in the last decades. Often compromises with respect to computational load,…

Systems and Control · Electrical Eng. & Systems 2024-02-21 Daniel D. Leister , Justin P. Koeln