English
Related papers

Related papers: Semiparametric count data regression for self-repo…

200 papers

Subjective wellness data can provide important information on the well-being of athletes and be used to maximize player performance and detect and prevent against injury. Wellness data, which are often ordinal and multivariate, include…

Applications · Statistics 2020-05-20 Erin M. Schliep , Toryn L. J. Schafer , Matthew Hawkey

The restricted mean survival time (RMST) model has been garnering attention as a way to provide a clinically intuitive measure: the mean survival time. RMST models, which use methods based on pseudo time-to-event values and inverse…

Methodology · Statistics 2024-06-11 Keisuke Hanada , Masahiro Kojima

There is increasing interest in learning how human brain networks vary as a function of a continuous trait, but flexible and efficient procedures to accomplish this goal are limited. We develop a Bayesian semiparametric model, which…

Methodology · Statistics 2017-02-02 Lu Wang , Daniele Durante , Rex E. Jung , David B. Dunson

This work presents a novel semi-supervised learning approach for data-driven modeling of asset failures when health status is only partially known in historical data. We combine a generative model parameterized by deep neural networks with…

Machine Learning · Computer Science 2017-09-05 Andre S. Yoon , Taehoon Lee , Yongsub Lim , Deokwoo Jung , Philgyun Kang , Dongwon Kim , Keuntae Park , Yongjin Choi

Introduction: Accounting for missing data by imputing or weighting conditional on covariates relies on the variable with missingness being observed at least some of the time for all unique covariate values. This requirement is referred to…

Applications · Statistics 2025-10-16 Paul N Zivich , Bonnie E Shook-Sa , Stephen R Cole , Eric T Lofgren , Jessie K Edwards

Nonresponse after probability sampling is a universal challenge in survey sampling, often necessitating adjustments to mitigate sampling and selection bias simultaneously. This study explored the removal of bias and effective utilization of…

Methodology · Statistics 2025-11-13 Kosuke Morikawa , Kenji Beppu , Wataru Aida

Public health data are often spatially dependent, but standard spatial regression methods can suffer from bias and invalid inference when the independent variable is associated with spatially-correlated residuals. This could occur if, for…

Methodology · Statistics 2025-04-10 Nate Wiecha , Jane A. Hoppin , Brian J. Reich

Imbalanced regression refers to prediction tasks where the target variable is skewed. This skewness hinders machine learning models, especially neural networks, which concentrate on dense regions and therefore perform poorly on…

Machine Learning · Computer Science 2025-08-11 Shayan Alahyari , Mike Domaratzki

The cause of failure in cohort studies that involve competing risks is frequently incompletely observed. To address this, several methods have been proposed for the semiparametric proportional cause-specific hazards model under a missing at…

Methodology · Statistics 2020-02-24 Giorgos Bakoyannis , Ying Zhang , Constantin T. Yiannoutsos

Statistical arbitrage strategies, such as pairs trading and its generalizations, rely on the construction of mean-reverting spreads enjoying a certain degree of predictability. Gaussian linear state-space processes have recently been…

Statistical Finance · Quantitative Finance 2009-05-19 Kostas Triantafyllopoulos , Giovanni Montana

We propose a novel personalized concept for the optimal treatment selection for a situation where the response is a multivariate vector, that could contain right-censored variables such as survival time. The proposed method can be applied…

Methodology · Statistics 2022-10-03 Chathura Siriwardhana , K. B. Kulasekera , Somnath Datta

Modern recording techniques enable neuroscientists to simultaneously study neural activity across large populations of neurons, with capturing predictor-dependent correlations being a fundamental challenge in neuroscience. Moreover, the…

Applications · Statistics 2025-02-04 Ganchao Wei

The article develops marginal models for multivariate longitudinal responses. Overall, the model consists of five regression submodels, one for the mean and four for the covariance matrix, with the latter resulting by considering various…

Methodology · Statistics 2020-12-18 Georgios Papageorgiou

Table retrieval is the task of retrieving the most relevant tables from large-scale corpora given natural language queries. However, structural and semantic discrepancies between unstructured text and structured tables make embedding…

Information Retrieval · Computer Science 2026-01-23 Shui-Hsiang Hsu , Tsung-Hsiang Chou , Chen-Jui Yu , Yao-Chung Fan

Prediction model training is often hindered by limited access to individual-level data due to privacy concerns and logistical challenges, particularly in biomedical research. Resampling-based self-training presents a promising approach for…

Methodology · Statistics 2025-03-18 Buxin Su , Jiaoyang Huang , Jin Jin , Bingxin Zhao

Large-scale population-level datasets, such as the UK Biobank and the All of Us Research Program, often lack covariates needed for a specific analysis, such as genetic or lifestyle measures, while related studies measure them. This creates…

Methodology · Statistics 2026-05-07 Huali Zhao , Tianying Wang

With social media communities increasingly becoming places where suicidal individuals post and congregate, natural language processing presents an exciting avenue for the development of automated suicide risk assessment systems. However,…

Computation and Language · Computer Science 2024-12-17 Max Lovitt , Haotian Ma , Song Wang , Yifan Peng

Handling imbalanced target distributions in regression poses a persistent challenge, as the underrepresentation of relevant target values can significantly hinder model performance. Existing data-level solutions often adapt…

Machine Learning · Computer Science 2026-03-12 António Pedro Pinheiro , Rita P. Ribeiro

A semi-parametric, non-linear regression model in the presence of latent variables is introduced. These latent variables can correspond to unmodeled phenomena or unmeasured agents in a complex networked system. This new formulation allows…

Machine Learning · Statistics 2018-06-29 Jonathan Mei , José M. F. Moura

The availability of mobile technologies has enabled the efficient collection prospective longitudinal, ecologically valid self-reported mood data from psychiatric patients. These data streams have potential for improving the efficiency and…

Applications · Statistics 2020-07-09 Yue Wu , Terry J. Lyons , Kate E. A. Saunders
‹ Prev 1 4 5 6 7 8 10 Next ›