中文
相关论文

相关论文: Modeling, dependence, classification, united stati…

200 篇论文

Boosting algorithms to simultaneously estimate and select predictor effects in statistical models have gained substantial interest during the last decade. This review article aims to highlight recent methodological developments regarding…

统计方法学 · 统计学 2014-11-19 Andreas Mayr , Harald Binder , Olaf Gefeller , Matthias Schmid

Supervised machine learning and predictive models have achieved an impressive standard today, enabling us to answer questions that were inconceivable a few years ago. Besides these successes, it becomes clear, that beyond pure prediction,…

Mutual information (MI) is a general measure of statistical dependence with widespread application across the sciences. However, estimating MI between multi-dimensional variables is challenging because the number of samples necessary to…

定量方法 · 定量生物学 2025-03-06 Gokul Gowri , Xiao-Kang Lun , Allon M. Klein , Peng Yin

In the era of big data, analysts usually explore various statistical models or machine learning methods for observed data in order to facilitate scientific discoveries or gain predictive power. Whatever data and fitting procedures are…

机器学习 · 统计学 2018-10-24 Jie Ding , Vahid Tarokh , Yuhong Yang

Modern longitudinal data, for example from wearable devices, measures biological signals on a fixed set of participants at a diverging number of time points. Traditional statistical methods are not equipped to handle the computational…

统计方法学 · 统计学 2023-03-23 Lan Luo , Jingshen Wang , Emily C. Hector

Nonparametric density estimation is an unsupervised learning problem. In this work we propose a two-step procedure that casts the density estimation problem in the first step into a supervised regression problem. The advantage is that we…

统计理论 · 数学 2024-06-04 Thijs Bos , Johannes Schmidt-Hieber

In the classical herding model, asymptotic learning refers to situations where individuals eventually take the correct action regardless of their private information. Classical results identify classes of information structures for which…

计算机科学与博弈论 · 计算机科学 2020-02-14 Itay Kavaler

Causal mediation analyses investigate the mechanisms through which causes exert their effects, and are therefore central to scientific progress. The literature on the non-parametric definition and identification of mediational effects in…

机器学习 · 统计学 2025-06-13 Richard Liu , Nicholas T. Williams , Kara E. Rudolph , Iván Díaz

Researchers in the behavioral and social sciences use linear discriminant analysis (LDA) for predictions of group membership (classification) and for identifying the variables most relevant to group separation among a set of continuous…

统计方法学 · 统计学 2025-05-28 Ricarda Graf , Marina Zeldovich , Sarah Friedrich

We review the most important statistical ideas of the past half century, which we categorize as: counterfactual causal inference, bootstrapping and simulation-based inference, overparameterized models and regularization, Bayesian multilevel…

统计方法学 · 统计学 2021-06-04 Andrew Gelman , Aki Vehtari

Cities are characterized by the presence of a dense population with a high potential for interactions between individuals of diverse backgrounds. They appear in parallel to the Neolithic revolution a few millennia ago. The advantages…

物理与社会 · 物理学 2022-10-12 Elsa Arcaute , Jose J. Ramasco

Schelling's model of segregation, first described in 1969, has become one of the best known models of self-organising behaviour. While Schelling's explicit concern was to understand the mechanisms underlying racial segregation in large…

多智能体系统 · 计算机科学 2016-09-12 George Barmpalias , Richard Elwes , Andy Lewis-Pye

We introduce a general modeling framework to predict the outcomes, at the population level, of individual psychology and behavior. The framework prescribes that researchers build a cost function that embodies knowledge of what trait values…

物理与社会 · 物理学 2007-05-23 Pierluigi Contucci , Stefano Ghirlanda

Random forests are a scheme proposed by Leo Breiman in the 2000's for building a predictor ensemble with a set of decision trees that grow in randomly selected subspaces of data. Despite growing interest and practical use, there has been…

机器学习 · 统计学 2012-03-28 Gérard Biau

The random forest algorithm, proposed by L. Breiman in 2001, has been extremely successful as a general-purpose classification and regression method. The approach, which combines several randomized decision trees and aggregates their…

统计理论 · 数学 2015-11-19 Gérard Biau , Erwan Scornet

Numerous recent studies have shown that Large Language Models (LLMs) are biased towards a Western and Anglo-centric worldview, which compromises their usefulness in non-Western cultural settings. However, "culture" is a complex,…

计算机与社会 · 计算机科学 2025-02-17 Sougata Saha , Saurabh Kumar Pandey , Monojit Choudhury

This paper is about how we study statistical methods. As an example, it uses the random regressions model, in which the intercept and slope of cluster-specific regression lines are modeled as a bivariate random effect. Maximizing this…

其他统计学 · 统计学 2019-05-22 James S. Hodges

Species sampling processes have long served as the fundamental framework for modeling random discrete distributions and exchangeable sequences. However, data arising from distinct but related sources require a broader notion of…

统计理论 · 数学 2026-02-03 Beatrice Franzolini , Antonio Lijoi , Igor Prünster , Giovanni Rebaudo

Big data are data on a massive scale in terms of volume, intensity, and complexity that exceed the capacity of standard software tools. They present opportunities as well as challenges to statisticians. The role of computational…

统计计算 · 统计学 2018-06-13 Chun Wang , Ming-Hui Chen , Elizabeth Schifano , Jing Wu , Jun Yan

Statistics is a uniquely difficult field to convey to the uninitiated. It sits astride the abstract and the concrete, the theoretical and the applied. It has a mathematical flavor and yet it is not simply a branch of mathematics. Its core…

应用统计 · 统计学 2011-01-06 Michael I. Jordan