中文
相关论文

相关论文: CUSUM ARL - Conditional or Unconditional?

200 篇论文

We propose Cartography Active Learning (CAL), a novel Active Learning (AL) algorithm that exploits the behavior of the model on individual instances during training as a proxy to find the most informative instances for labeling. CAL is…

计算与语言 · 计算机科学 2022-05-10 Mike Zhang , Barbara Plank

Active Learning (AL) is increasingly important in a broad range of applications. Two main AL principles to obtain accurate classification with few labeled data are refinement of the current decision boundary and exploration of poorly…

机器学习 · 计算机科学 2012-10-19 Jens Roeder , Boaz Nadler , Kevin Kunzmann , Fred A. Hamprecht

The paper is about detecting changes in the parameters of certain parameterized stochastic models. We apply CUSUM (Cumulated Sums) type test statistics that are based on martingale difference sequences.

统计理论 · 数学 2014-07-22 Fanni Nedényi

Reinforcement learning (RL) has become a pivotal component of large language model (LLM) post-training, and agentic RL extends this paradigm to operate as agents through multi-turn interaction and tool use. Scaling such systems exposes two…

分布式、并行与集群计算 · 计算机科学 2025-10-08 Zheyue Tan , Mustapha Abdullahi , Tuo Shi , Huining Yuan , Zelai Xu , Chao Yu , Boxun Li , Bo Zhao

Linear ARCH (LARCH) processes were introduced by Robinson [J. Econometrics 47 (1991) 67--84] to model long-range dependence in volatility and leverage. Basic theoretical properties of LARCH processes have been investigated in the recent…

统计理论 · 数学 2010-01-13 Jan Beran , Martin Schützner

We present general principles underlying analysis of the dependence of random variables (outputs) on deterministic conditions (inputs). Random outputs recorded under mutually exclusive input values are labeled by these values and considered…

量子物理 · 物理学 2015-01-27 Ehtibar N. Dzhafarov , Janne V. Kujala

Single-task RL agents are typically trained under a fixed reward function, which limits their robustness to reward misspecification and their ability to adapt to changing preferences. We introduce Reward-Conditioned Reinforcement Learning…

机器学习 · 计算机科学 2026-05-20 Michal Nauman , Marek Cygan , Pieter Abbeel

A weakly dependent time series regression model with multivariate covariates and univariate observations is considered, for which we develop a procedure to detect whether the nonparametric conditional mean function is stable in time against…

统计理论 · 数学 2019-01-25 Maria Mohr , Natalie Neumeyer

We present new estimators for the statistical analysis of the dependence of the mean gap time length between consecutive recurrent events, on a set of explanatory random variables and in the presence of right censoring. The dependence is…

应用统计 · 统计学 2021-09-10 Ioana Schiopu-Kratina , Hai Yan Liu , Mayer Alvo , Pierre-Jerome Bergeron

Continual reinforcement learning (continual RL) seeks to formalize the notions of lifelong learning and endless adaptation in RL. In particular, the aim of continual RL is to develop RL agents that can maintain a careful balance between…

机器学习 · 计算机科学 2026-05-05 Juan Sebastian Rojas , Chi-Guhn Lee

We study the limiting behavior of continuous time trawl processes which are defined using an infinitely divisible random measure of a time dependent set. In this way one is able to define separately the marginal distribution and the…

概率论 · 数学 2017-08-10 Danijel Grahovac , Nikolai N. Leonenko , Murad S. Taqqu

Autonomous operations of robots in unknown environments are challenging due to the lack of knowledge of the dynamics of the interactions, such as the objects' movability. This work introduces a novel Causal Reinforcement Learning approach…

Predictive models trained on observational data often fail to generalise to the distributions they encounter when deployed, especially when the training data is a product of the system being optimised. Recommender systems are a canonical…

机器学习 · 统计学 2026-05-27 Yorgos Felekis , Michael O'Riordan , Oriol Corcoll , Ciarán M. Gilligan-Lee

Training of elite athletes requires regular physiological and medical monitoring to plan the schedule, intensity and volume of training, and subsequent recovery. In sports medicine, ECG-based analyses are well established. However, they…

应用统计 · 统计学 2018-09-27 Marcel Młyńczak , Hubert Krysztofiak

Continuous Time Random Maxima (CTRM) are a generalization of classical extreme value theory: Instead of observing random events at regular intervals in time, the waiting times between the events are also random variables with arbitrary…

概率论 · 数学 2017-02-02 Katharina Hees , Hans-Peter Scheffler

Autonomous vehicles (AVs) make driving decisions without humans, making dependability assurance critical. Scenario-based testing is widely used to evaluate AVs under diverse conditions, with reinforcement learning (RL) generating test…

软件工程 · 计算机科学 2026-04-29 Jiahui Wu , Chengjie Lu , Aitor Arrieta , Shaukat Ali

Since the advent of autonomous driving technology, it has experienced remarkable progress over the last decade. However, most existing research still struggles to address the challenges posed by environments where multiple vehicles have to…

多智能体系统 · 计算机科学 2025-08-01 Jing Wang , Yan Jin , Fei Ding , Chongfeng Wei

An intermittent nonlinear map generating subdiffusion is investigated. Computer simulations show that the generalized diffusion coefficient of this map has a fractal, discontinuous dependence on control parameters. An amended continuous…

The usual development of the continuous time random walk (CTRW) assumes that jumps and time intervals are a two-dimensional set of independent and identically distributed random variables. In this paper we address the theoretical setting of…

数据分析、统计与概率 · 物理学 2008-09-29 Miquel Montero , Jaume Masoliver

Instead of randomly acquiring training data points, Uncertainty-based Active Learning (UAL) operates by querying the label(s) of pivotal samples from an unlabeled pool selected based on the prediction uncertainty, thereby aiming at…

机器学习 · 计算机科学 2024-08-27 Amir Hossein Rahmati , Mingzhou Fan , Ruida Zhou , Nathan M. Urban , Byung-Jun Yoon , Xiaoning Qian