中文
相关论文

相关论文: Nested Dirichlet Process For Population Size Estim…

200 篇论文

This study presents a semi-nonparametric Latent Class Choice Model (LCCM) with a flexible class membership component. The proposed model formulates the latent classes using mixture models as an alternative approach to the traditional random…

计量经济学 · 经济学 2023-08-07 Georges Sfeir , Maya Abou-Zeid , Filipe Rodrigues , Francisco Camara Pereira , Isam Kaysi

In ecology, the description of species composition and biodiversity calls for statistical methods that involve estimating features of interest in unobserved samples based on an observed one. In the last decade, the Bayesian nonparametrics…

统计方法学 · 统计学 2026-04-28 Alessandro Colombi , Raffaele Argiento , Federico Camerlenghi , Lucia Paci

The latent position cluster model is a popular model for the statistical analysis of network data. This approach assumes that there is an underlying latent space in which the actors follow a finite mixture distribution. Moreover, actors…

统计计算 · 统计学 2013-08-23 Nial Friel , Caitriona Ryan , Jason Wyse

We propose a Multi-Task Learning (MTL) paradigm based deep neural network architecture, called MTCNet (Multi-Task Crowd Network) for crowd density and count estimation. Crowd count estimation is challenging due to the non-uniform scale…

机器学习 · 计算机科学 2025-04-16 Abhay Kumar , Nishant Jain , Suraj Tripathi , Chirag Singh , Kamal Krishna

Multidimensional network data can have different levels of complexity, as nodes may be characterized by heterogeneous individual-specific features, which may vary across the networks. This paper introduces a class of models for…

统计方法学 · 统计学 2021-04-01 Silvia D'Angelo , Marco Alfò , Thomas Brendan Murphy

In this paper, we consider capture-recapture experiments with heterogenous catchability. In the setting we consider, the widespread Huggins-Alho estimator is not very suitable and we introduce and study a new generalized Horvitz-Thompson…

应用统计 · 统计学 2011-11-09 Yakir Berchenko , Richard G. White , Cyprian Wejnert , Simon D. W. Frost

Item response theory (IRT) models typically rely on a normality assumption for subject-specific latent traits, which is often unrealistic in practice. Semiparametric extensions based on Dirichlet process mixtures offer a more flexible…

The Manifold Hypothesis is a widely accepted tenet of Machine Learning which asserts that nominally high-dimensional data are in fact concentrated near a low-dimensional manifold, embedded in high-dimensional space. This phenomenon is…

统计方法学 · 统计学 2025-03-24 Nick Whiteley , Annie Gray , Patrick Rubin-Delanchy

The use of high-dimensional data for targeted therapeutic interventions requires new ways to characterize the heterogeneity observed across subgroups of a specific population. In particular, models for partially exchangeable data are needed…

统计方法学 · 统计学 2020-08-18 Francesco Denti , Federico Camerlenghi , Michele Guindani , Antonietta Mira

This paper introduces a flexible time-varying network vector autoregressive model framework for large-scale time series. A latent group structure is imposed on the heterogeneous and node-specific time-varying momentum and network spillover…

统计方法学 · 统计学 2024-03-12 Degui Li , Bin Peng , Songqiao Tang , Weibiao Wu

We consider the problem of estimating the number of distinct elements in a large data set (or, equivalently, the support size of the distribution induced by the data set) from a random sample of its elements. The problem occurs in many…

机器学习 · 计算机科学 2021-06-17 Talya Eden , Piotr Indyk , Shyam Narayanan , Ronitt Rubinfeld , Sandeep Silwal , Tal Wagner

Survival analysis provides a well-established framework for modeling time-to-event data, with hazard and survival functions formally defined as population-level quantities. In applied work, however, these quantities are often interpreted as…

统计方法学 · 统计学 2026-03-26 Xijia Liu

We initiate the study of a new model of supervised learning under privacy constraints. Imagine a medical study where a dataset is sampled from a population of both healthy and unhealthy individuals. Suppose healthy individuals have no…

机器学习 · 计算机科学 2020-08-04 Raef Bassily , Shay Moran , Anupama Nandi

Although Recurrent Neural Network (RNN) has been a powerful tool for modeling sequential data, its performance is inadequate when processing sequences with multiple patterns. In this paper, we address this challenge by introducing a novel…

机器学习 · 计算机科学 2019-02-28 Kui Zhao , Yuechuan Li , Chi Zhang , Cheng Yang , Huan Xu

Many datasets describing contacts in a population suffer from incompleteness due to population sampling and underreporting of contacts. Data-driven simulations of spreading processes using such incomplete data lead to an underestimation of…

物理与社会 · 物理学 2017-09-07 Julie Fournet , Alain Barrat

Hierarchical probabilistic models, such as mixture models, are used for cluster analysis. These models have two types of variables: observable and latent. In cluster analysis, the latent variable is estimated, and it is expected that…

机器学习 · 统计学 2017-06-26 Keisuke Yamazaki

Accurate inference on population dynamics, such as migration and changes in population size, is essential for policymaking, resource allocation and demographic research. Traditional censuses are expensive, infrequent and not timely, leading…

应用统计 · 统计学 2026-03-27 Lucy Y Brown , Eleni Matechou , Bruno Santos , Eleonora Mussino

Sequential importance sampling algorithms have been defined to estimate likelihoods in models of ancestral population processes. However, these algorithms are based on features of the models with constant population size, and become…

统计理论 · 数学 2016-03-24 Coralie Merle , Raphaël Leblois , François Rousset , Pierre Pudlo

In this paper, we introduce the \textit{Layer-Peeled Model}, a nonconvex yet analytically tractable optimization program, in a quest to better understand deep neural networks that are trained for a sufficiently long time. As the name…

机器学习 · 计算机科学 2022-05-11 Cong Fang , Hangfeng He , Qi Long , Weijie J. Su

This paper targets the question of predicting machine learning classification model performance, when taking into account the number of training examples per class and not just the overall number of training examples. This leads to the a…

机器学习 · 计算机科学 2024-03-12 Thomas Mühlenstädt , Jelena Frtunikj