中文
相关论文

相关论文: Model-based Bootstrap of Controlled Markov Chains

200 篇论文

Rough volatility models have recently been empirically shown to provide a good fit to historical volatility time series and implied volatility smiles of SPX options. They are continuous-time stochastic volatility models, whose volatility…

数理金融 · 定量金融 2021-11-01 Jingtang Ma , Wensheng Yang , Zhenyu Cui

This study investigates a data-driven machine learning approach to predict membrane fouling in critically ill patients undergoing Continuous Renal Replacement Therapy (CRRT). Using time-series data from an ICU, 16 clinically selected…

Robust model predictive control (MPC) is a well-known control technique for model-based control with constraints and uncertainties. In classic robust tube-based MPC approaches, an open-loop control sequence is computed via periodically…

系统与控制 · 电气工程与系统科学 2022-06-13 Xinglong Zhang , Jiahang Liu , Xin Xu , Shuyou Yu , Hong Chen

In Change point detection task Likelihood Ratio Test (LRT) is sequentially applied in a sliding window procedure. Its high values indicate changes of parametric distribution in the data sequence. Correspondingly LRT values require…

统计理论 · 数学 2017-10-23 Nazar Buzun , Valeriy Avanesov

We propose distribution-free runs-based control charts for detecting location shifts. Using the fact that given the number of total successes, the outcomes of a sequence of Bernoulli trials are random permutations, we are able to control…

统计方法学 · 统计学 2025-11-19 Tung-Lung Wu

Reliability is an important tool for evaluating the performance of modern networks. Currently, it is NP-hard and #P-hard to calculate the exact reliability of a binary-state network when the reliability of each component is assumed to be…

系统与控制 · 电气工程与系统科学 2022-02-17 Wei-Chang Yeh

Clinical prediction models are increasingly used to support patient care, yet many deep learning-based approaches remain unstable, as their predictions can vary substantially when trained on different samples from the same population. Such…

机器学习 · 计算机科学 2026-02-13 Sara Matijevic , Christopher Yau

Recent developments in Reinforcement learning have significantly enhanced sequential decision-making in uncertain environments. Despite their strong performance guarantees, most existing work has focused primarily on improving the…

统计理论 · 数学 2025-08-13 Bo Pan , Jianya Lu , Yafei Wang , Hao Li , Bei Jiang , Linglong Kong

Offline reinforcement learning (RL) suffers from the distribution shift between the offline dataset and the online environment. In multi-agent RL (MARL), this distribution shift may arise from the nonstationary opponents in the online…

机器学习 · 计算机科学 2025-02-25 Tao Li , Juan Guevara , Xinhong Xie , Quanyan Zhu

Balancing exploration and exploitation is crucial in reinforcement learning (RL). In this paper, we study model-based posterior sampling for reinforcement learning (PSRL) in continuous state-action spaces theoretically and empirically.…

机器学习 · 计算机科学 2021-11-18 Ying Fan , Yifei Ming

This paper investigates the (in)-consistency of various bootstrap methods for making inference on a change-point in time in the Cox model with right censored survival data. A criterion is established for the consistency of any bootstrap…

统计方法学 · 统计学 2013-08-01 Gongjun Xu , Bodhisattva Sen , Zhiliang Ying

Continuous-time reinforcement learning (CTRL) provides a natural framework for sequential decision-making in dynamic environments where interactions evolve continuously over time. While CTRL has shown growing empirical success, its ability…

机器学习 · 计算机科学 2025-12-04 Runze Zhao , Yue Yu , Ruhan Wang , Chunfeng Huang , Dongruo Zhou

The paper studies an improved estimate for the rate of convergence for nonlinear homogeneous discrete-time Markov chains. These processes are nonlinear in terms of the distribution law. Hence, the transition kernels are dependent on the…

概率论 · 数学 2021-05-21 Aleksandr Shchegolev

In the framework of Model Predictive Control (MPC), the control input is typically computed by solving optimization problems repeatedly online. For general nonlinear systems, the online optimization problems are non-convex and…

最优化与控制 · 数学 2021-03-31 Zheming Wang , Raphaël M. Jungers

The existing theory of penalized quantile regression for longitudinal data has focused primarily on point estimation. In this work, we investigate statistical inference. We propose a wild residual bootstrap procedure and show that it is…

计量经济学 · 经济学 2022-05-10 Carlos Lamarche , Thomas Parker

Automating complex industrial robots requires precise nonlinear control and efficient energy management. This paper introduces a data-driven nonlinear model predictive control (NMPC) framework to optimize control under multiple objectives.…

机器人学 · 计算机科学 2024-11-22 Dexian Ma , Bo Zhou

This work provides a state-of-the-art survey of continual safe online reinforcement learning (COSRL) methods. We discuss theoretical aspects, challenges, and open questions in building continual online safe reinforcement learning…

机器学习 · 计算机科学 2026-01-09 Timofey Tomashevskiy

We propose a novel Markov chain Monte-Carlo (MCMC) method for reverse engineering the topological structure of stochastic reaction networks, a notoriously challenging problem that is relevant in many modern areas of research, like…

统计方法学 · 统计学 2018-10-08 Daniel F. Linder , Grzegorz A. Rempala

We develop algorithms with low regret for learning episodic Markov decision processes based on kernel approximation techniques. The algorithms are based on both the Upper Confidence Bound (UCB) as well as Posterior or Thompson Sampling…

机器学习 · 计算机科学 2019-11-06 Sayak Ray Chowdhury , Aditya Gopalan

This paper introduces a data-based integral sliding mode control scheme for robustification of model-reference controllers, accommodating generic multivariable linear systems with unknown dynamics and affected by matched disturbances.…

系统与控制 · 电气工程与系统科学 2026-02-06 Giorgio Riva , Gian Paolo Incremona , Simone Formentin , Antonella Ferrara