English
Related papers

Related papers: A copula-based boosting model for time-to-event pr…

200 papers

Traditional statistical and machine learning methods typically assume that the training and test data follow the same distribution. However, this assumption is frequently violated in real-world applications, where the training data in the…

Methodology · Statistics 2025-07-08 Hanxuan Ye , Hongzhe Li

Copulas provide an attractive approach for constructing multivariate distributions with flexible marginal distributions and different forms of dependences. Of particular importance in many areas is the possibility of explicitly forecasting…

Methodology · Statistics 2018-05-22 Feng Li , Yanfei Kang

Treatment policy estimands are frequently favored by regulators, as they assess the effect of treatment assignment regardless of post-randomization events. Despite best efforts, missing data due to study discontinuation cannot be fully…

Methodology · Statistics 2026-05-13 Ajmal Oodally , Craig Wang , Zheng Li , Tim Morris , Tobias Mütze , Arunava Chakravartty

This paper is concerned with modeling the dependence structure of two (or more) time-series in the presence of a (possible multivariate) covariate which may include past values of the time series. We assume that the covariate influences…

Statistics Theory · Mathematics 2018-12-11 Natalie Neumeyer , Marek Omelka , Sarka Hudecova

In this paper, we considered the problem of dependent censoring models with a positive probability that the times of failure are equal. In this context, we proposed to consider the Marshall-Olkin type model and studied some properties of…

Statistics Theory · Mathematics 2023-09-08 Mikael Escobar-Bach , Salima Helali

Latent Gaussian models and boosting are widely used techniques in statistics and machine learning. Tree-boosting shows excellent prediction accuracy on many data sets, but potential drawbacks are that it assumes conditional independence of…

Machine Learning · Computer Science 2022-08-24 Fabio Sigrist

Motivated by challenges in the analysis of biomedical data and observational studies, we develop statistical boosting for the general class of bivariate distributional copula regression with arbitrary marginal distributions, which is suited…

Methodology · Statistics 2024-03-05 Guillermo Briseño Sanchez , Nadja Klein , Hannah Klinkhammer , Andreas Mayr

Recent deep clustering models have produced impressive clustering performance. However, a common issue with existing methods is the disparity between global and local feature structures. While local structures typically show strong…

Computer Vision and Pattern Recognition · Computer Science 2025-11-27 Hanyang Li , Yuheng Jia , Hui Liu , Junhui Hou

This paper considers the problem of inferring the causal effect of a variable $Z$ on a dependently censored survival time $T$. We allow for unobserved confounding variables, such that the error term of the regression model for $T$ is…

Statistics Theory · Mathematics 2024-10-02 Gilles Crommen , Jad Beyhum , Ingrid Van Keilegom

We present a new variable selection method based on model-based gradient boosting and randomly permuted variables. Model-based boosting is a tool to fit a statistical model while performing variable selection at the same time. A drawback of…

Machine Learning · Statistics 2017-02-16 Janek Thomas , Tobias Hepp , Andreas Mayr , Bernd Bischl

We introduce a copula mixture model to perform dependency-seeking clustering when co-occurring samples from different data sources are available. The model takes advantage of the great flexibility offered by the copulas framework to extend…

Methodology · Statistics 2012-07-03 Melanie Rey , Volker Roth

With insurers benefiting from ever-larger amounts of data of increasing complexity, we explore a data-driven method to model dependence within multilevel claims in this paper. More specifically, we start from a non-parametric estimator for…

Methodology · Statistics 2024-01-17 Marie Michaelides , Hélène Cossette , Mathieu Pigeon

Boosting methods are widely used in statistical learning to deal with high-dimensional data due to their variable selection feature. However, those methods lack straightforward ways to construct estimators for the precision of the…

Methodology · Statistics 2021-06-10 Boyao Zhang , Colin Griesbach , Cora Kim , Nadia Müller-Voggel , Elisabeth Bergherr

Stress-strength models are widely used to assess the reliability of systems under uncertain conditions. While most studies assume independence between stress and strength variables, such an assumption may be unrealistic in many practical…

Methodology · Statistics 2026-04-15 Fatih Kızılaslan

Multi-agent imitation learning aims to train multiple agents to perform tasks from demonstrations by learning a mapping between observations and actions, which is essential for understanding physical, social, and team-play systems. However,…

Machine Learning · Computer Science 2021-07-13 Hongwei Wang , Lantao Yu , Zhangjie Cao , Stefano Ermon

Existing survival analysis techniques heavily rely on strong modelling assumptions and are, therefore, prone to model misspecification errors. In this paper, we develop an inferential method based on ideas from conformal prediction, which…

Methodology · Statistics 2023-04-25 Emmanuel J. Candès , Lihua Lei , Zhimei Ren

Conventional survival metrics, such as Harrell's concordance index (CI) and the Brier Score, rely on the independent censoring assumption for valid inference with right-censored data. However, in the presence of so-called dependent…

Machine Learning · Statistics 2025-05-20 Christian Marius Lillelund , Shi-ang Qi , Russell Greiner

A robust model for time series forecasting is highly important in many domains, including but not limited to financial forecast, air temperature and electricity consumption. To improve forecasting performance, traditional approaches usually…

Machine Learning · Computer Science 2019-09-19 Long H. Nguyen , Zhenhe Pan , Opeyemi Openiyi , Hashim Abu-gellban , Mahdi Moghadasi , Fang Jin

We consider the problem of classification in a comparison-based setting: given a set of objects, we only have access to triplet comparisons of the form "object $x_i$ is closer to object $x_j$ than to object $x_k$." In this paper we…

Machine Learning · Statistics 2019-05-30 Michaël Perrot , Ulrike von Luxburg

Survival analysis is a valuable tool for estimating the time until specific events, such as death or cancer recurrence, based on baseline observations. This is particularly useful in healthcare to prognostically predict clinically important…

Machine Learning · Computer Science 2024-01-11 Ahmed H. Shahin , An Zhao , Alexander C. Whitehead , Daniel C. Alexander , Joseph Jacob , David Barber