English
Related papers

Related papers: Individual Data Protected Integrative Regression A…

200 papers

Susceptible-Exposed-Infectious-Recovered (SEIR) models with inter-individual variation in susceptibility or exposure to infection were proposed early in the COVID-19 pandemic as a potential element of the mathematical/statistical toolset…

Applications · Statistics 2025-10-28 Ibrahim Mohammed , Chris Robertson , M. Gabriela M. Gomes

This paper presents a histogram based reversible data hiding (RDH) scheme, which divides image pixels into different cell frequency bands to sort them for data embedding. Data hiding is more efficient in lower cell frequency bands because…

Image and Video Processing · Electrical Eng. & Systems 2020-10-19 Ammar Mohammadi , Mansour Nakhkash

The debiased estimator is a crucial tool in statistical inference for high-dimensional model parameters. However, constructing such an estimator involves estimating the high-dimensional inverse Hessian matrix, incurring significant…

Machine Learning · Statistics 2023-12-18 Jiyuan Tu , Weidong Liu , Xiaojun Mao , Mingyue Xu

Competing risk analysis considers event times due to multiple causes, or of more than one event types. Commonly used regression models for such data include 1) cause-specific hazards model, which focuses on modeling one type of event while…

Applications · Statistics 2017-04-27 Jiayi Hou , Anthony Paravati , Ronghui Xu , James Murphy

Variable selection in ultra-high dimensional regression problems has become an important issue. In such situations, penalized regression models may face computational problems and some pre screening of the variables may be necessary. A…

Methodology · Statistics 2020-05-01 Abhik Ghosh , Magne Thoresen

Personalized health analytics increasingly rely on population benchmarks to provide contextual insights such as ''How do I compare to others like me?'' However, cohort-based aggregation of health data introduces nontrivial privacy risks,…

Cryptography and Security · Computer Science 2026-01-21 Richik Chakraborty , Lawrence Liu , Syed Hasnain

Integrated IPD-AD analysis, which combines individual participant data (IPD) with aggregate data (AD), is increasingly recognized as an effective strategy for generating more reliable and generalizable inferences from heterogeneous studies.…

Methodology · Statistics 2026-03-03 Ming-Yueh Huang , Jing Qin , Chiung-Yu Huang

Repeated-measure designs allow comparisons within a group as well as between groups, and are commonly referred to as split-plot designs. While originating in agricultural experiments, they are now widely used in medical research,…

Computation · Statistics 2025-12-22 Paavo Sattler , Nils Hichert

Sufficient dimension reduction (SDR) is a popular tool in regression analysis, which replaces the original predictors with a minimal set of their linear combinations. However, the estimated linear combinations generally contain all original…

Computation · Statistics 2020-12-16 Lei Yan , Xin Chen

Economics and social science research often require analyzing datasets of sensitive personal information at fine granularity, with models fit to small subsets of the data. Unfortunately, such fine-grained analysis can easily reveal…

Machine Learning · Computer Science 2020-07-13 Daniel Alabi , Audra McMillan , Jayshree Sarathy , Adam Smith , Salil Vadhan

High-dimensional data often arise from clinical genomics research to infer relevant predictors of a particular trait. A way to improve the predictive performance is to include information on the predictors derived from prior knowledge or…

Methodology · Statistics 2023-03-13 Claudio Busatto , Mark van de Wiel

As one of the most fundamental problems in machine learning, statistics and differential privacy, Differentially Private Stochastic Convex Optimization (DP-SCO) has been extensively studied in recent years. However, most of the previous…

Machine Learning · Computer Science 2021-08-10 Lijie Hu , Shuo Ni , Hanshen Xiao , Di Wang

Analysis of a probabilistic system often requires to learn the joint probability distribution of its random variables. The computation of the exact distribution is usually an exhaustive precise analysis on all executions of the system. To…

Information Theory · Computer Science 2023-07-19 Fabrizio Biondi , Yusuke Kawamoto , Axel Legay , Louis-Marie Traonouez

Private information retrieval (PIR) is a database query protocol that provides user privacy, in that the user can learn a particular entry of the database of his interest but his query would be hidden from the data centre. Symmetric private…

Quantum Physics · Physics 2021-01-19 Wen Yu Kon , Charles Ci Wen Lim

Combining an internal individual-level study with readily available external summary statistics promises major efficiency gains at minimal additional cost, yet heterogeneity between sources can bias estimates for the internal target…

Methodology · Statistics 2026-01-29 Kosuke Morikawa , Sho Komukai , Satoshi Hattori

While machine learning has become pervasive in as diversified fields as industry, healthcare, social networks, privacy concerns regarding the training data have gained a critical importance. In settings where several parties wish to…

Cryptography and Security · Computer Science 2023-04-07 Arnaud Grivet Sébert , Martin Zuber , Oana Stan , Renaud Sirdey , Cédric Gouy-Pailler

We propose a distributed quadratic inference function framework to jointly estimate regression parameters from multiple potentially heterogeneous data sources with correlated vector outcomes. The primary goal of this joint integrative…

Methodology · Statistics 2022-07-28 Emily C. Hector , Peter X. -K. Song

In this work, a simple and efficient dual iterative refinement (DIR) method is proposed for dense correspondence between two nearly isometric shapes. The key idea is to use dual information, such as spatial and spectral, or local and global…

Computer Vision and Pattern Recognition · Computer Science 2020-11-20 Rui Xiang , Rongjie Lai , Hongkai Zhao

Synthesizing information from multiple data sources is critical to ensure knowledge generalizability. Integrative analysis of multi-source data is challenging due to the heterogeneity across sources and data-sharing constraints due to…

Methodology · Statistics 2023-01-03 Zijian Guo , Xiudi Li , Larry Han , Tianxi Cai

With the increasing adoption of electronic health records, there is an increasing interest in developing individualized treatment rules, which recommend treatments according to patients' characteristics, from large observational data.…

Methodology · Statistics 2021-05-05 Muxuan Liang , Young-Geun Choi , Yang Ning , Maureen A Smith , Ying-Qi Zhao