English
Related papers

Related papers: Model-based Bootstrap of Controlled Markov Chains

200 papers

Rough volatility models have recently been empirically shown to provide a good fit to historical volatility time series and implied volatility smiles of SPX options. They are continuous-time stochastic volatility models, whose volatility…

Mathematical Finance · Quantitative Finance 2021-11-01 Jingtang Ma , Wensheng Yang , Zhenyu Cui

This study investigates a data-driven machine learning approach to predict membrane fouling in critically ill patients undergoing Continuous Renal Replacement Therapy (CRRT). Using time-series data from an ICU, 16 clinically selected…

Robust model predictive control (MPC) is a well-known control technique for model-based control with constraints and uncertainties. In classic robust tube-based MPC approaches, an open-loop control sequence is computed via periodically…

Systems and Control · Electrical Eng. & Systems 2022-06-13 Xinglong Zhang , Jiahang Liu , Xin Xu , Shuyou Yu , Hong Chen

In Change point detection task Likelihood Ratio Test (LRT) is sequentially applied in a sliding window procedure. Its high values indicate changes of parametric distribution in the data sequence. Correspondingly LRT values require…

Statistics Theory · Mathematics 2017-10-23 Nazar Buzun , Valeriy Avanesov

We propose distribution-free runs-based control charts for detecting location shifts. Using the fact that given the number of total successes, the outcomes of a sequence of Bernoulli trials are random permutations, we are able to control…

Methodology · Statistics 2025-11-19 Tung-Lung Wu

Reliability is an important tool for evaluating the performance of modern networks. Currently, it is NP-hard and #P-hard to calculate the exact reliability of a binary-state network when the reliability of each component is assumed to be…

Systems and Control · Electrical Eng. & Systems 2022-02-17 Wei-Chang Yeh

Clinical prediction models are increasingly used to support patient care, yet many deep learning-based approaches remain unstable, as their predictions can vary substantially when trained on different samples from the same population. Such…

Machine Learning · Computer Science 2026-02-13 Sara Matijevic , Christopher Yau

Recent developments in Reinforcement learning have significantly enhanced sequential decision-making in uncertain environments. Despite their strong performance guarantees, most existing work has focused primarily on improving the…

Statistics Theory · Mathematics 2025-08-13 Bo Pan , Jianya Lu , Yafei Wang , Hao Li , Bei Jiang , Linglong Kong

Offline reinforcement learning (RL) suffers from the distribution shift between the offline dataset and the online environment. In multi-agent RL (MARL), this distribution shift may arise from the nonstationary opponents in the online…

Machine Learning · Computer Science 2025-02-25 Tao Li , Juan Guevara , Xinhong Xie , Quanyan Zhu

Balancing exploration and exploitation is crucial in reinforcement learning (RL). In this paper, we study model-based posterior sampling for reinforcement learning (PSRL) in continuous state-action spaces theoretically and empirically.…

Machine Learning · Computer Science 2021-11-18 Ying Fan , Yifei Ming

This paper investigates the (in)-consistency of various bootstrap methods for making inference on a change-point in time in the Cox model with right censored survival data. A criterion is established for the consistency of any bootstrap…

Methodology · Statistics 2013-08-01 Gongjun Xu , Bodhisattva Sen , Zhiliang Ying

Continuous-time reinforcement learning (CTRL) provides a natural framework for sequential decision-making in dynamic environments where interactions evolve continuously over time. While CTRL has shown growing empirical success, its ability…

Machine Learning · Computer Science 2025-12-04 Runze Zhao , Yue Yu , Ruhan Wang , Chunfeng Huang , Dongruo Zhou

The paper studies an improved estimate for the rate of convergence for nonlinear homogeneous discrete-time Markov chains. These processes are nonlinear in terms of the distribution law. Hence, the transition kernels are dependent on the…

Probability · Mathematics 2021-05-21 Aleksandr Shchegolev

In the framework of Model Predictive Control (MPC), the control input is typically computed by solving optimization problems repeatedly online. For general nonlinear systems, the online optimization problems are non-convex and…

Optimization and Control · Mathematics 2021-03-31 Zheming Wang , Raphaël M. Jungers

The existing theory of penalized quantile regression for longitudinal data has focused primarily on point estimation. In this work, we investigate statistical inference. We propose a wild residual bootstrap procedure and show that it is…

Econometrics · Economics 2022-05-10 Carlos Lamarche , Thomas Parker

Automating complex industrial robots requires precise nonlinear control and efficient energy management. This paper introduces a data-driven nonlinear model predictive control (NMPC) framework to optimize control under multiple objectives.…

Robotics · Computer Science 2024-11-22 Dexian Ma , Bo Zhou

This work provides a state-of-the-art survey of continual safe online reinforcement learning (COSRL) methods. We discuss theoretical aspects, challenges, and open questions in building continual online safe reinforcement learning…

Machine Learning · Computer Science 2026-01-09 Timofey Tomashevskiy

We propose a novel Markov chain Monte-Carlo (MCMC) method for reverse engineering the topological structure of stochastic reaction networks, a notoriously challenging problem that is relevant in many modern areas of research, like…

Methodology · Statistics 2018-10-08 Daniel F. Linder , Grzegorz A. Rempala

We develop algorithms with low regret for learning episodic Markov decision processes based on kernel approximation techniques. The algorithms are based on both the Upper Confidence Bound (UCB) as well as Posterior or Thompson Sampling…

Machine Learning · Computer Science 2019-11-06 Sayak Ray Chowdhury , Aditya Gopalan

This paper introduces a data-based integral sliding mode control scheme for robustification of model-reference controllers, accommodating generic multivariable linear systems with unknown dynamics and affected by matched disturbances.…

Systems and Control · Electrical Eng. & Systems 2026-02-06 Giorgio Riva , Gian Paolo Incremona , Simone Formentin , Antonella Ferrara