English
Related papers

Related papers: Model-free screening procedure for ultrahigh-dimen…

200 papers

The purpose of this paper is to construct confidence intervals for the regression coefficients in the Fine-Gray model for competing risks data with random censoring, where the number of covariates can be larger than the sample size. Despite…

Methodology · Statistics 2019-04-10 Jue Hou , Jelena Bradic , Ronghui Xu

The paper presents new metrics to quantify and test for (i) the equality of distributions and (ii) the independence between two high-dimensional random vectors. We show that the energy distance based on the usual Euclidean distance cannot…

Methodology · Statistics 2019-10-01 Shubhadeep Chakraborty , Xianyang Zhang

We develop a framework for post model selection inference, via marginal screening, in linear regression. At the core of this framework is a result that characterizes the exact distribution of linear functions of the response $y$,…

Methodology · Statistics 2014-03-03 Jason D Lee , Jonathan E Taylor

Two popular variable screening methods under the ultra-high dimensional setting with the desirable sure screening property are the sure independence screening (SIS) and the forward regression (FR). Both are classical variable screening…

Methodology · Statistics 2015-11-05 Ming-Yen Cheng , Sanying Feng , Gaorong Li , Heng Lian

Huntington disease (HD) is a neurodegenerative disease with progressively worsening symptoms. Accurately modeling time to HD diagnosis is essential for clinical trial design. Langbehn's model, the CAG-Age Product (CAP) model, the Prognostic…

In this paper, we present a new algorithm for semi-supervised representation learning. In this algorithm, we first find a vector representation for the labels of the data points based on their local positions in the space. Then, we map the…

Machine Learning · Computer Science 2020-08-05 Ershad Banijamali , Ali Ghodsi

Although the independent censoring assumption is commonly used in survival analysis, it can be violated when the censoring time is related to the survival time, which often happens in many practical applications. To address this issue, we…

Methodology · Statistics 2024-08-28 Huazhen Yu , Lixin Zhang

Model selection is crucial to high-dimensional learning and inference for contemporary big data applications in pinpointing the best set of covariates among a sequence of candidate interpretable models. Most existing work assumes implicitly…

Methodology · Statistics 2018-03-21 Emre Demirkaya , Yang Feng , Pallavi Basu , Jinchi Lv

In this paper, we present the general theory of embedding independence tests on Hilbert spaces that generalizes the concepts of distance covariance, distance multivariance and HSIC. This is done by defining new types of kernel on an $n$…

Functional Analysis · Mathematics 2024-11-14 Jean Carlo Guella

Measurements of systems taken along a continuous functional dimension, such as time or space, are ubiquitous in many fields, from the physical and biological sciences to economics and engineering.Such measurements can be viewed as…

Maximum mean discrepancy (MMD), also called energy distance or N-distance in statistics and Hilbert-Schmidt independence criterion (HSIC), specifically distance covariance in statistics, are among the most popular and successful approaches…

Machine Learning · Statistics 2018-08-03 Zoltan Szabo , Bharath K. Sriperumbudur

Density estimation is a classical problem in statistics and has received considerable attention when both the data has been fully observed and in the case of partially observed (censored) samples. In survival analysis or clinical trials, a…

Applications · Statistics 2018-04-18 German A. Schnaidt Grez , Brani Vidakovic

We introduce two novel non-parametric statistical hypothesis tests. The first test, called the relative test of dependency, enables us to determine whether one source variable is significantly more dependent on a first target variable or a…

Artificial Intelligence · Computer Science 2016-11-18 Wacha Bounliphone , Eugene Belilovsky , Arthur Tenenhaus , Ioannis Antonoglou , Arthur Gretton , Matthew B. Blashcko

We propose an iterative variable selection scheme for high-dimensional data with binary outcomes. The scheme adopts a structured screen-and-select framework and uses non-local prior-based Bayesian model selection within the same. The…

Methodology · Statistics 2022-11-08 Nilotpal Sanyal

Motivation: A branching processes model yields an unevenly stochastically distributed dataset that consists of sparse and dense regions. This work addresses the problem of precisely evaluating parameters for such a model. Applying a…

Machine Learning · Statistics 2023-05-09 Iurii S. Nagornov

In the context of the usual calibration model, we consider the case in which the independent variable is unobservable, but a pre-fixed value on its surrogate is available. Thus, considering controlled variables and assuming that the…

Applications · Statistics 2008-02-06 Betsabé G. Blas Achic , Mônica C. Sandoval , Olga Satomi Yoshida

Modern bio-technologies have produced a vast amount of high-throughput data with the number of predictors far greater than the sample size. In order to identify more novel biomarkers and understand biological mechanisms, it is vital to…

Machine Learning · Statistics 2018-05-18 Kevin He , Jian Kang , Hyokyoung Grace Hong , Ji Zhu , Yanming Li , Huazhen Lin , Han Xu , Yi Li

We discuss the development of reliability acceptance sampling plans under progressive Type-I interval censoring schemes in the presence of competing causes of failure. We consider a general framework to accommodate the presence of…

Applications · Statistics 2025-01-22 Rathin Das , Soumya Roy , Biswabrata Pradhan

Renal cell carcinoma represents a significant global health challenge with a low survival rate. This research aimed to devise a comprehensive deep-learning model capable of predicting survival probabilities in patients with renal cell…

Computer Vision and Pattern Recognition · Computer Science 2023-07-10 Maryamalsadat Mahootiha , Hemin Ali Qadir , Jacob Bergsland , Ilangko Balasingham

We propose a new class of metrics, called the survival independence divergence (SID), to test dependence between a right-censored outcome and covariates. A key technique for deriving the SIDs is to use a counting process strategy, which…

Methodology · Statistics 2026-05-06 Jinhong Li , Jicai Liu , Jinhong You , Riquan Zhang