English
Related papers

Related papers: Comparing Spatial Regression to Random Forests for…

200 papers

Battery performance datasets are typically non-normal and multicollinear. Extrapolating such datasets for model predictions needs attention to such characteristics. This study explores the impact of data normality in building machine…

Machine Learning · Computer Science 2021-11-05 Shovan Chowdhury , Yuxiao Lin , Boryann Liaw , Leslie Kerby

Plant biomass estimation is critical due to the variability of different environmental factors and crop management practices associated with it. The assessment is largely impacted by the accurate prediction of different environmental…

Artificial Intelligence · Computer Science 2023-02-07 Syeda Nyma Ferdous , Xin Li , Kamalakanta Sahoo , Richard Bergman

The last decade has seen an explosion in data sources available for the monitoring and prediction of environmental phenomena. While several inferential methods have been developed that make predictions on the underlying process by combining…

Methodology · Statistics 2023-03-06 Eun-Hye Yoo , Andrew Zammit-Mangion , Michael G. Chipeta

Remote sensing data are increasingly available and frequently used to produce forest attributes maps. The sampling strategy of the calibration plots may directly affect predictions and map qualities. The aim of this manuscript is to…

Applications · Statistics 2024-08-09 Andrey Ramirez Luigui , Jean-Pierre Renaud , Cédric Vega

In recent years, the growing availability of biomedical datasets featuring numerous longitudinal covariates has motivated the development of several multi-step methods for the dynamic prediction of survival outcomes. These methods employ…

Methodology · Statistics 2026-01-14 Mirko Signorelli , Sophie Retif

Random forests are a statistical learning method widely used in many areas of scientific research because of its ability to learn complex relationships between input and output variables and also its capacity to handle high-dimensional…

Machine Learning · Statistics 2024-02-19 Louis Capitaine , Jérémie Bigot , Rodolphe Thiébaut , Robin Genuer

Spatial processes observed in various fields, such as climate and environmental science, often occur on a large scale and demonstrate spatial nonstationarity. Fitting a Gaussian process with a nonstationary Mat\'ern covariance is…

Machine Learning · Statistics 2023-06-21 Pratik Nag , Yiping Hong , Sameh Abdulah , Ghulam A. Qadir , Marc G. Genton , Ying Sun

Spatial patterning is common in ecological systems and has been extensively studied via different modeling approaches. Individual-based models (IBMs) accurately describe nonlinear interactions at the organism level and the stochastic…

Populations and Evolution · Quantitative Biology 2025-04-16 Anudeep Surendran , David Pinto-Ramos , Rafael Menezes , Ricardo Martinez-Garcia

We propose a method for variable selection in the intensity function of spatial point processes that combines sparsity-promoting estimation with noise-robust model selection. As high-resolution spatial data becomes increasingly available…

Methodology · Statistics 2025-10-30 Dominik Sturm , Ivo F. Sbalzarini

The problem of subgroups is ubiquitous in scientific research (ex. disease heterogeneity, spatial distributions in ecology...), and piecewise regression is one way to deal with this phenomenon. Morse-Smale regression offers a way to…

Machine Learning · Statistics 2017-08-22 Colleen M. Farrelly

A compositional tree refers to a tree structure on a set of random variables where each random variable is a node and composition occurs at each non-leaf node of the tree. As a generalization of compositional data, compositional trees…

Methodology · Statistics 2021-04-20 Bingkai Wang , Brian S. Caffo , Xi Luo , Chin-Fu Liu , Andreia V. Faria , Michael I. Miller , Yi Zhao

A random forest prediction can be computed by the scalar product of the labels of the training examples and a set of weights that are determined by the leafs of the forest into which the test object falls; each prediction can hence be…

Machine Learning · Computer Science 2023-11-27 Henrik Boström

Random forests have become popular for clinical risk prediction modelling. In a case study on predicting ovarian malignancy, we observed training c-statistics close to 1. Although this suggests overfitting, performance was competitive on…

In modelling time series data coming from different sources, frequencies can easily vary since some variable can be measured at higher frequencies, others, at lower frequencies. Given data measured over spatial units and at varying…

Methodology · Statistics 2025-03-05 Vladimir A. Malabanan , Joseph Ryan G. Lansangan , Erniel B. Barrios

Spatial confounding between the spatial random effects and fixed effects covariates has been recently discovered and showed that it may bring misleading interpretation to the model results. Solutions to alleviate this problem are based on…

Methodology · Statistics 2016-05-17 Marcos O. Prates , Erica C. Rodrigues , Renato M. Assunção

Random forest (Leo Breiman 2001a) (RF) is a non-parametric statistical method requiring no distributional assumptions on covariate relation to the response. RF is a robust, nonlinear technique that optimizes predictive accuracy by fitting…

Computation · Statistics 2016-12-30 John Ehrlinger

Understanding historical forest dynamics, specifically changes in forest biomass and carbon stocks, has become critical for assessing current forest climate benefits and projecting future benefits under various policy, regulatory, and…

Applications · Statistics 2023-08-29 Lucas K. Johnson , Michael J. Mahoney , Madeleine L. Desrochers , Colin M. Beier

Models of spatial transition probabilities, or equivalently, transiogram models have been recently proposed as spatial continuity measures in categorical fields. In this paper, properties of transiogram models are examined analytically, and…

Applications · Statistics 2016-06-08 Guofeng Cao , Phaedon Kyriakidis , Michael Goodchild

A key challenge in estimating causal effects from observational data is handling confounding and is commonly achieved through weighting methods that balance distribution of covariates between treatment and control groups. Weighting…

Methodology · Statistics 2025-12-23 Simion De , Jared D. Huling

Multi-omics data, that is, datasets containing different types of high-dimensional molecular variables (often in addition to classical clinical variables), are increasingly generated for the investigation of various diseases. Nevertheless,…

Machine Learning · Statistics 2020-12-22 Moritz Herrmann , Philipp Probst , Roman Hornung , Vindi Jurinovic , Anne-Laure Boulesteix
‹ Prev 1 8 9 10 Next ›