English
Related papers

Related papers: Cautionary note on "Semiparametric modeling of gro…

200 papers

We develop inference procedures for longitudinal data where some of the measurements are censored by fixed constants. We consider a semi-parametric quantile regression model that makes no distributional assumptions. Our research is…

Statistics Theory · Mathematics 2009-04-02 Huixia Judy Wang , Mendel Fygenson

A targeted learning (TL) framework is developed to estimate the difference in the restricted mean survival time (RMST) for a clinical trial with time-to-event outcomes. The approach starts by defining the target estimand as the RMST…

Methodology · Statistics 2026-01-21 Man Jin , Yixin Fang

Mixture of autoregressions (MoAR) models provide a model-based approach to the clustering of time series data. The maximum likelihood (ML) estimation of MoAR models requires the evaluation of products of large numbers of densities of normal…

Computation · Statistics 2016-10-19 Hien D Nguyen , Geoffrey J McLachlan , Pierre Orban , Pierre Bellec , Andrew L Janke

In a longitudinal study, measures of key variables might be incomplete or partially recorded due to drop-out, loss to follow-up, or early termination of the study occurring before the advent of the event of interest. In this paper, we focus…

Methodology · Statistics 2020-08-19 Roland A. Matsouaka , Folefac D. Atem

This paper addresses the problem of identifying and estimating the causal effect of a treatment in the presence of unmeasured confounding and various types of right-censoring. Examples of these censoring mechanisms are administrative…

Statistics Theory · Mathematics 2025-03-19 Ilias Willems , Sara Rutten , Gilles Crommen , Ingrid Van Keilegom

In this article, we propose some new generalizations of M-estimation procedures for single-index regression models in presence of randomly right-censored responses. We derive consistency and asymptotic normality of our estimates. The…

Statistics Theory · Mathematics 2008-12-18 Olivier Lopez

Although many fairness criteria have been proposed to ensure that machine learning algorithms do not exhibit or amplify our existing social biases, these algorithms are trained on datasets that can themselves be statistically biased. In…

Machine Learning · Computer Science 2023-05-04 Yiqiao Liao , Parinaz Naghizadeh

In this article, the analysis of left truncated and right censored competing risks data is carried out, under the assumption of the latent failure times model. It is assumed that there are two competing causes of failures, although most of…

Methodology · Statistics 2020-08-19 Debasis Kundu , Debanjan Mitra , Ayon Ganguly

Let P represent the source population with complete data, containing covariate $\mathbf{Z}$ and response $T$, and Q the target population, where only the covariate $\mathbf{Z}$ is available. We consider a setting with both label shift and…

Methodology · Statistics 2025-06-27 Yuxiang Zong , Yanyuan Ma , Ingrid Van Keilegom

We present a unified parametric framework for modal regression applicable to continuous positive distributions, with explicit support for right-censored observations. The key contribution is a systematic analytical reparameterization of…

Methodology · Statistics 2026-03-10 Christian E. Galarza , Víctor H. Lachos

The benefits and capabilities of pre-trained language models (LLMs) in current and future innovations are vital to any society. However, introducing and using LLMs comes with biases and discrimination, resulting in concerns about equality,…

Computers and Society · Computer Science 2023-12-05 Vithya Yogarajan , Gillian Dobbie , Te Taka Keegan , Rostam J. Neuwirth

Double (debiased) machine learning (DML) has seen widespread use in recent years for learning causal/structural parameters, in part due to its flexibility and adaptability to high-dimensional nuisance functions as well as its ability to…

Methodology · Statistics 2024-09-12 Abhinandan Dalal , Patrick Blöbaum , Shiva Kasiviswanathan , Aaditya Ramdas

Adequate sampling space coverage is the keystone to effectively train trustworthy Machine Learning models. Unfortunately, real data do carry several inherent risks due to the many potential biases they exhibit when gathered without a proper…

Machine Learning · Computer Science 2025-03-27 Antonio Maratea , Rita Perna

Modern language models are trained on large amounts of data. These data inevitably include controversial and stereotypical content, which contains all sorts of biases related to gender, origin, age, etc. As a result, the models express…

Computation and Language · Computer Science 2025-09-03 Aleksandra Sorokovikova , Pavel Chizhov , Iuliia Eremenko , Ivan P. Yamshchikov

Parameter estimates in misspecified models converge to pseudo-true parameter values, which minimize a population objective function. Pseudo-true values often differ from quantities of economic interest, raising questions of how, if at all,…

Econometrics · Economics 2026-04-20 Isaiah Andrews , Harvey Barnhard , Jacob Carlson

Interval-censored data, in which the event time is only known to lie in some time interval, arise commonly in practice; for example, in a medical study in which patients visit clinics or hospitals at pre-scheduled times, and the events of…

Methodology · Statistics 2017-07-21 Wei Fu , Jeffrey S. Simonoff

Continuous-time multi-state survival models can be used to describe health-related processes over time. In the presence of interval-censored times for transitions between the living states, the likelihood is constructed using transition…

Methodology · Statistics 2017-03-24 Robson J. M. Machado , Ardo van den Hout

Accurate time-to-event prediction is integral to decision-making, informing medical guidelines, hiring decisions, and resource allocation. Survival analysis, the quantitative framework used to model time-to-event data, accounts for patients…

Machine Learning · Computer Science 2025-08-08 Vincent Jeanselme , Brian Tom , Jessica Barrett

In this review, we present a simple guide for researchers to obtain pseudo-random samples with censored data. We focus our attention on the most common types of censored data, such as type I, type II, and random censoring. We discussed the…

When data are collected subject to a detection limit, observations below the detection limit may be considered censored. In addition, the domain of such observations may be restricted; for example, values may be required to be non-negative.…

Applications · Statistics 2020-06-30 Justin R. Williams , Hyung-Woo Kim , Catherine M. Crespi