中文
相关论文

相关论文: Modelling Preference Data with the Wallenius Distr…

200 篇论文

Statistical modelling of sports data has become more and more popular in the recent years and different types of models have been proposed to achieve a variety of objectives: from identifying the key characteristics which lead a team to win…

应用统计 · 统计学 2019-11-21 Andrea Gabrio

Stacking is a widely used model averaging technique that asymptotically yields optimal predictions among linear averages. We show that stacking is most effective when model predictive performance is heterogeneous in inputs, and we can…

统计方法学 · 统计学 2021-10-29 Yuling Yao , Gregor Pirš , Aki Vehtari , Andrew Gelman

The ratio between the probability that two distributions $R$ and $P$ give to points $x$ are known as importance weights or propensity scores and play a fundamental role in many different fields, most notably, statistics and machine…

机器学习 · 计算机科学 2021-03-11 Parikshit Gopalan , Omer Reingold , Vatsal Sharan , Udi Wieder

In Bayesian theory, calculating a posterior probability distribution is highly important but usually difficult. Therefore, some methods have been put forward to deal with such problem, among which, the most popular one is the asymptotic…

统计方法学 · 统计学 2012-07-20 Zai-Ying Zhou

Ranking or assessing centrality in multivariate and non-Euclidean data is difficult because there is no canonical order and many depth notions become computationally fragile in high-dimensional or structured settings. We introduce a…

统计方法学 · 统计学 2026-02-24 Lingfeng Lyu , Doudou Zhou

Combining distributions is an important issue in decision theory and Bayesian inference. Logarithmic pooling is a popular method to aggregate expert opinions by using a set of weights that reflect the reliability of each information source.…

There is an innate human tendency, one might call it the "league table mentality," to construct rankings. Schools, hospitals, sports teams, movies, and myriad other objects are ranked even though their inherent multi-dimensionality would…

计量经济学 · 经济学 2021-09-16 Jiaying Gu , Roger Koenker

Multidimensional indexes are ubiquitous, and popular, but present non-negligible normative choices when it comes to attributing weights to their dimensions. This paper provides a more rigorous approach to the choice of weights by defining a…

计量经济学 · 经济学 2025-04-09 Lidia Ceriani , Chiara Gigliarano , Paolo Verme

It is now practically the norm for data to be very high dimensional in areas such as genetics, machine vision, image analysis and many others. When analyzing such data, parametric models are often too inflexible while nonparametric…

统计方法学 · 统计学 2011-05-31 Abhishek Bhattacharya , Garritt Page , David Dunson

A central statistical goal is to choose between alternative explanatory models of data. In many modern applications, such as population genetics, it is not possible to apply standard methods based on evaluating the likelihood functions of…

统计计算 · 统计学 2013-02-25 Dennis Prangle , Paul Fearnhead , Murray P. Cox , Patrick J. Biggs , Nigel P. French

Percentiles and more generally, quantiles are commonly used in various contexts to summarize data. For most distributions, there is exactly one quantile that is unbiased. For distributions like the Gaussian that have the same mean and…

统计方法学 · 统计学 2022-01-11 Rohit Pandey

The goal of this presentation is to build an efficient non-parametric Bayes classifier in the presence of large numbers of predictors. When analyzing such data, parametric models are often too inflexible while non-parametric procedures tend…

统计方法学 · 统计学 2013-01-07 Abhishek Bhattacharya

Datasets are rarely a realistic approximation of the target population. Say, prevalence is misrepresented, image quality is above clinical standards, etc. This mismatch is known as sampling bias. Sampling biases are a major hindrance for…

In this paper we consider a well-known generalization of the Barab\'asi and Albert preferential attachment model - the Buckley-Osthus model. Buckley and Osthus proved that in this model the degree sequence has a power law distribution. As a…

Distribution data refers to a data set where each sample is represented as a probability distribution, a subject area receiving burgeoning interest in the field of statistics. Although several studies have developed…

统计方法学 · 统计学 2024-02-09 Ryo Okano , Masaaki Imaizumi

Often the rows (cases, objects) of a dataset have weights. For instance, the weight of a case may reflect the number of times it has been observed, or its reliability. For analyzing such data many rowwise weighted techniques are available,…

统计计算 · 统计学 2024-07-08 Peter J. Rousseeuw

Ranking and comparing items is crucial for collecting information about preferences in many areas, from marketing to politics. The Mallows rank model is among the most successful approaches to analyse rank data, but its computational…

统计方法学 · 统计学 2017-04-28 Valeria Vitelli , Øystein Sørensen , Marta Crispino , Arnoldo Frigessi , Elja Arjas

The elicitation of an ordinal judgment on multiple alternatives is often required in many psychological and behavioral experiments to investigate preference/choice orientation of a specific population. The Plackett-Luce model is one of the…

统计方法学 · 统计学 2016-10-10 Cristina Mollica , Luca Tardella

Many policies allocate harms or benefits that are uncertain in nature: they produce distributions over the population in which individuals have different probabilities of incurring harm or benefit. Comparing different policies thus involves…

计算机与社会 · 计算机科学 2021-03-11 Hoda Heidari , Solon Barocas , Jon Kleinberg , Karen Levy

Model averaging is a useful and robust method for dealing with model uncertainty in statistical analysis. Often, it is useful to consider data subset selection at the same time, in which model selection criteria are used to compare models…

统计方法学 · 统计学 2023-10-26 Ethan T. Neil , Jacob W. Sitison