English
Related papers

Related papers: sensobol: an R package to compute variance-based s…

200 papers

Biomarker data is often subject to limits of quantification and/or limits of detection. Statistically, this corresponds to left- or interval-censoring. To be able to associate a censored time-to-event endpoint to a biomarker covariate, the…

Computation · Statistics 2014-04-29 Stanislas Hubeaux , Kaspar Rufibach

In this paper we describe the implementation of semi-structured deep distributional regression, a flexible framework to learn conditional distributions based on the combination of additive regression models and deep networks. Our…

Instrumental variables regression is a tool that is commonly used in the analysis of observational data. The instrumental variables are used to make causal inference about the effect of a certain exposure in the presence of unmeasured…

Methodology · Statistics 2023-09-07 Valentin Vancak , Arvid Sjölander

We introduce varbvs, a suite of functions written in R and MATLAB for regression analysis of large-scale data sets using Bayesian variable selection methods. We have developed numerical optimization algorithms based on variational…

Computation · Statistics 2017-09-21 Peter Carbonetto , Xiang Zhou , Matthew Stephens

Bayesian synthetic likelihood (BSL) is a popular method for estimating the parameter posterior distribution for complex statistical models and stochastic processes that possess a computationally intractable likelihood function. Instead of…

Computation · Statistics 2019-07-26 Ziwen An , Leah F South , Christopher Drovandi

The Stochastic Volatility (SV) model and its variants are widely used in the financial sector while recurrent neural network (RNN) models are successfully used in many large-scale industrial applications of Deep Learning. Our article…

Econometrics · Economics 2022-01-25 Trong-Nghia Nguyen , Minh-Ngoc Tran , David Gunawan , R. Kohn

SimOmics is an R package designed to generate realistic, multivariate, and multi-omics synthetic datasets. It is intended for use in benchmarking, method development, and reproducibility in bioinformatics, particularly in the context of…

Genomics · Quantitative Biology 2025-07-15 Kaitao Lai

The PAWN index is gaining traction among the modelling community as a sensitivity measure. However, the robustness to its design parameters has not yet been scrutinized: the size ($N$) and sampling ($\varepsilon$) of the model output, the…

Applications · Statistics 2020-09-03 Arnald Puy , Samuele Lo Piano , Andrea Saltelli

We propose a new statistical estimation framework for a large family of global sensitivity analysis indices. Our approach is based on rank statistics and uses an empirical correlation coefficient recently introduced by Chatterjee [9]. We…

Methodology · Statistics 2026-05-25 Fabrice Gamboa , Pierre Gremaud , Thierry Klein , Agnès Lagnoux

Variance-based global sensitivity analysis (GSA) can provide a wealth of information when applied to complex models. A well-known Achilles' heel of this approach is its computational cost which often renders it unfeasible in practice. An…

Numerical Analysis · Mathematics 2026-01-08 John Darges , Alen Alexanderian , Pierre Gremaud

Variance-based sensitivity indices have established themselves as a reference among practitioners of sensitivity analysis of model outputs. A variance-based sensitivity analysis typically produces the first-order sensitivity indices $S_j$…

Applications · Statistics 2022-03-02 Samuele Lo Piano , Federico Ferretti , Arnald Puy , Daniel Albrecht , Andrea Saltelli

For models evaluated at a random set of independent variables, the variance-based Shapley effects range between Sobol' indices, and the corresponding total indices admit derivative-based upper-bounds. Such relationships fail when the inputs…

Statistics Theory · Mathematics 2026-05-28 Matieyendou Lamboni

Researchers would often like to leverage data from a collection of sources (e.g., primary studies in a meta-analysis) to estimate causal effects in a target population of interest. However, traditional meta-analytic methods do not produce…

Methodology · Statistics 2025-05-15 Guanbo Wang , Sean McGrath , Yi Lian

Gaussian processes (GPs) are well-known tools for modeling dependent data with applications in spatial statistics, time series analysis, or econometrics. In this article, we present the R package varycoef that implements estimation,…

Computation · Statistics 2021-06-07 Jakob A. Dambon , Fabio Sigrist , Reinhard Furrer

Variable importance measures are the main tools to analyze the black-box mechanisms of random forests. Although the mean decrease accuracy (MDA) is widely accepted as the most efficient variable importance measure for random forests, little…

Machine Learning · Statistics 2022-03-02 Clément Bénard , Sébastien da Veiga , Erwan Scornet

The R add-on package FDboost is a flexible toolbox for the estimation of functional regression models by model-based boosting. It provides the possibility to fit regression models for scalar and functional response with effects of scalar as…

Computation · Statistics 2018-04-27 Sarah Brockhaus , David Rügamer , Sonja Greven

In the field of computer experiments sensitivity analysis aims at quantifying the relative importance of each input parameter (or combinations thereof) of a computational model with respect to the model output uncertainty. Variance…

Computation · Statistics 2014-05-23 Bruno Sudret , Chu Van Mai

CensSpatial is an R package for analyzing spatial censored data through linear models. It offers a set of tools for simulating, estimating, making predictions, and performing local influence diagnostics for outlier detection. The package…

Methodology · Statistics 2021-10-13 Jose A. Ordonez , Christian E. Galarza , Victor H. Lachos

A new unimodal distribution family indexed by the mode and three other parameters is derived from a mixture of a Gumbel distribution for the maximum and a Gumbel distribution for the minimum. Properties of the proposed distribution are…

Methodology · Statistics 2024-07-02 Qingyang Liu , Xianzheng Huang , Haiming Zhou

Factor importance measures the impact of each feature on output prediction accuracy. Many existing works focus on the model-based importance, but an important feature in one learning algorithm may hold little significance in another model.…

Methodology · Statistics 2025-06-24 Chaofan Huang , V. Roshan Joseph