中文
相关论文

相关论文: Bayesian hierarchical modelling of sparse count pr…

200 篇论文

Multi-task learning models using Gaussian processes (GP) have been developed and successfully applied in various applications. The main difficulty with this approach is the computational cost of inference using the union of examples from…

机器学习 · 计算机科学 2012-11-29 Yuyang Wang , Roni Khardon

Crowdsourcing has emerged as an effective means for performing a number of machine learning tasks such as annotation and labelling of images and other data sets. In most early settings of crowdsourcing, the task involved classification,…

机器学习 · 计算机科学 2020-06-03 Desmond Cai , Duc Thien Nguyen , Shiau Hong Lim , Laura Wynter

We consider a novel Bayesian approach to estimation, uncertainty quantification, and variable selection for a high-dimensional linear regression model under sparsity. The number of predictors can be nearly exponentially large relative to…

统计方法学 · 统计学 2025-02-28 Samhita Pal , Subhashis Ghoshal

Learning product representations that reflect complementary relationship plays a central role in e-commerce recommender system. In the absence of the product relationships graph, which existing methods rely on, there is a need to detect the…

信息检索 · 计算机科学 2019-12-02 Da Xu , Chuanwei Ruan , Jason Cho , Evren Korpeoglu , Sushant Kumar , Kannan Achan

This article presents an approach to Bayesian semiparametric inference for Gaussian multivariate response regression. We are motivated by various small and medium dimensional problems from the physical and social sciences. The statistical…

统计方法学 · 统计学 2020-06-18 Georgios Papageorgiou , Benjamin C. Marshall

This article introduces methods for constructing prediction bounds or intervals for the number of future failures from heterogeneous reliability field data. We focus on within-sample prediction where early data from a failure-time process…

统计方法学 · 统计学 2021-04-13 Colin Lewis-Beck , Qinglong Tian , William Q. Meeker

Sparse Gaussian Processes are a key component of high-throughput Bayesian Optimisation (BO) loops; however, we show that existing methods for allocating their inducing points severely hamper optimisation performance. By exploiting the…

机器学习 · 计算机科学 2023-02-24 Henry B. Moss , Sebastian W. Ober , Victor Picheny

Hierarchical Bayesian methods enable information sharing across multiple related regression problems. While standard practice is to model regression parameters (effects) as (1) exchangeable across datasets and (2) correlated to differing…

统计方法学 · 统计学 2021-07-15 Brian L. Trippe , Hilary K. Finucane , Tamara Broderick

Despite the growing availability of sensing and data in general, we remain unable to fully characterise many in-service engineering systems and structures from a purely data-driven approach. The vast data and resources available to capture…

机器学习 · 计算机科学 2023-09-20 Elizabeth J Cross , Timothy J Rogers , Daniel J Pitchforth , Samuel J Gibson , Matthew R Jones

One of the focal points of the modern literature on Bayesian nonparametrics has been the problem of clustering, or partitioning, where each data point is modeled as being associated with one and only one of some collection of groups called…

统计理论 · 数学 2013-10-02 Tamara Broderick , Michael I. Jordan , Jim Pitman

In this work we review the application of the theory of Gaussian processes to the modeling of noise in pulsar-timing data analysis, and we derive various useful and optimized representations for the likelihood expressions that are needed in…

广义相对论与量子宇宙学 · 物理学 2014-11-19 Rutger van Haasteren , Michele Vallisneri

Classically, statistical datasets have a larger number of data points than features ($n > p$). The standard model of classical statistics caters for the case where data points are considered conditionally independent given the parameters.…

机器学习 · 统计学 2022-03-16 Sijia Li , Martín López-García , Neil D. Lawrence , Luisa Cutillo

Count data is prevalent in various fields like ecology, medical research, and genomics. In high-dimensional settings, where the number of features exceeds the sample size, feature selection becomes essential. While frequentist methods like…

统计方法学 · 统计学 2024-10-22 The Tien Mai

In genomics, differential abundance and expression analyses are complicated by the compositional nature of sequence count data, which reflect only relative-not absolute-abundances or expression levels. Many existing methods attempt to…

统计方法学 · 统计学 2025-12-16 Won Gu , Francesca Chiaromonte , Justin D. Silverman

Reliable demand forecasts are critical for the effective supply chain management. Several endogenous and exogenous variables can influence the dynamics of demand, and hence a single statistical model that only consists of historical sales…

应用统计 · 统计学 2019-09-09 Mahdi Abolghasemi , Ali Eshragh , Jason Hurley , Behnam Fahimnia

Analyzing demographic data collected across multiple populations, time periods, and age groups is challenging due to the interplay of high dimensionality, demographic heterogeneity among groups, and stochastic variability within smaller…

应用统计 · 统计学 2025-12-12 Gregor Zens

Clustering is one of the most widely used procedures in the analysis of microarray data, for example with the goal of discovering cancer subtypes based on observed heterogeneity of genetic marks between different tissues. It is well-known…

统计方法学 · 统计学 2009-04-21 Heng Lian

Raking is widely used in categorical data modeling and survey practice but faced with methodological and computational challenges. We develop a Bayesian paradigm for raking by incorporating the marginal constraints as a prior distribution…

统计方法学 · 统计学 2020-06-24 Yajuan Si , Peigen Zhou

Time series forecasting plays an increasingly important role in modern business decisions. In today's data-rich environment, people often aim to choose the optimal forecasting model for their data. However, identifying the optimal model…

应用统计 · 统计学 2021-12-17 Xixi Li , Fotios Petropoulos , Yanfei Kang

Hyperparameter tuning is a challenging problem especially when the system itself involves uncertainty. Due to noisy function evaluations, optimization under uncertainty can be computationally expensive. In this paper, we present a novel…

机器学习 · 计算机科学 2025-10-09 Akash Yadav , Ruda Zhang