English
Related papers

Related papers: Testing Differences Statistically with the Leiden …

200 papers

It is shown that under certain circumstances in particular for small datasets the recently proposed citation impact indicators I3(6PR) and R(6,k) behave inconsistently when additional papers or citations are taken into consideration. Three…

Applications · Statistics 2013-01-31 Michael Schreiber

Statistical NLP systems are frequently evaluated and compared on the basis of their performances on a single split of training and test data. Results obtained using a single split are, however, subject to sampling noise. In this paper we…

Computation and Language · Computer Science 2007-05-23 Yuval Krymolowski

Large language models (LLMs) frequently achieve impressive scores on standardized benchmarks, yet accuracy alone offers a limited view of their capabilities. Evaluating open-source LLMs through leaderboards faces persistent issues like data…

This paper considers ranking inference of $n$ items based on the observed data on the top choice among $M$ randomly selected items at each trial. This is a useful modification of the Plackett-Luce model for $M$-way ranking with only the top…

Methodology · Statistics 2023-01-09 Jianqing Fan , Zhipeng Lou , Weichen Wang , Mengxin Yu

The methods presented in this paper allow for a statistical analysis revealing centers of excellence around the world using programs that are freely available. Based on Web of Science data, field-specific excellence can be identified in…

Digital Libraries · Computer Science 2011-06-29 Lutz Bornmann , Loet Leydesdorff

Assessing research that pushes the boundaries of knowledge is challenging because such work is extremely infrequent, accounting for only about 0.01 per cent of all research outputs. Consequently, knowledge about how to evaluate this type of…

Digital Libraries · Computer Science 2026-05-19 Alonso Rodriguez-Navarro

A new test statistic based on success runs of weighted deviations is introduced. Its use for observations sampled from independent normal distributions is worked out in detail. It supplements the classic $\chi^{2}$ test which ignores the…

Statistics Theory · Mathematics 2017-04-10 Frederik Beaujean , Allen Caldwell

The 2017 Fake News Challenge Stage 1 (FNC-1) shared task addressed a stance classification task as a crucial first step towards detecting fake news. To date, there is no in-depth analysis paper to critically discuss FNC-1's experimental…

Citation metrics are the best tools for research assessments. However, current metrics may be misleading in research systems that pursue simultaneously different goals, such as the advance of science and incremental innovations, because…

Digital Libraries · Computer Science 2023-09-27 Alonso Rodriguez-Navarro , Ricardo Brito

This paper considers the problem of ranking objects based on their latent merits using data from pairwise interactions. We allow for incomplete observation of these interactions and study what can be inferred about rankings in such…

Econometrics · Economics 2025-09-23 Federico Crippa , Danil Fedchenko

The world's collective knowledge is evolving through research and new scientific discoveries. It is becoming increasingly difficult to objectively rank the impact research institutes have on global advancements. However, since the funding,…

Machine Learning · Computer Science 2020-12-25 Vlad Sandulescu , Mihai Chiru

Comparing alternatives in pairs is a well-known method of ranking creation. Experts are asked to perform a series of binary comparisons and then, using mathematical methods, the final ranking is prepared. As experts conduct the individual…

Discrete Mathematics · Computer Science 2018-12-12 Konrad Kułakowski

Ordered probit and logit models have been frequently used to estimate the mean ranking of happiness outcomes (and other ordinal data) across groups. However, it has been recently highlighted that such ranking may not be identified in most…

Econometrics · Economics 2022-06-08 Le-Yu Chen , Ekaterina Oparina , Nattavudh Powdthavee , Sorawoot Srisuma

Creating test collections for offline retrieval evaluation requires human effort to judge documents' relevance. This expensive activity motivated much work in developing methods for constructing benchmarks with fewer assessment costs. In…

Information Retrieval · Computer Science 2023-08-29 David Otero , Javier Parapar , Nicola Ferro

In reaction to a previous critique(Opthof & Leydesdorff, 2010), the Center for Science and Technology Studies (CWTS) in Leiden proposed to change their old "crown" indicator in citation analysis into a new one. Waltman et al. (2011)argue…

Digital Libraries · Computer Science 2011-02-16 Tobias Opthof , Loet Leydesdorff

We study the statistics of citations made to the indexed Science journals in the Journal Citation Reports during the period 2004-2013 using different measures. We consider different measures which quantify the impact of the journals. To our…

Digital Libraries · Computer Science 2018-09-20 Abdul Khaleque , Arnab Chatterjee , Parongama Sen

In this paper we extend the principle of proportional representation to rankings. We consider the setting where alternatives need to be ranked based on approval preferences. In this setting, proportional representation requires that…

Computer Science and Game Theory · Computer Science 2016-12-06 Piotr Skowron , Martin Lackner , Markus Brill , Dominik Peters , Edith Elkind

Studies to compare the survival of two or more groups using time-to-event data are of high importance in medical research. The gold standard is the log-rank test, which is optimal under proportional hazards. As the latter is no simple…

Methodology · Statistics 2022-10-25 Ina Dormuth , Tiantian Liu , Jin Xu , Markus Pauly , Marc Ditzhaus

PageRank has become a key element in the success of search engines, allowing to rank the most important hits in the top screen of results. One key aspect that distinguishes PageRank from other prestige measures such as in-degree is its…

Information Retrieval · Computer Science 2007-05-23 Santo Fortunato , Marian Boguna , Alessandro Flammini , Filippo Menczer

We discuss the paper "Citation Statistics" by the Joint Committee on Quantitative Assessment of Research [arXiv:0910.3529]. In particular, we focus on a necessary feature of "good" measures for ranking scientific authors: that good measures…

Methodology · Statistics 2009-10-20 Sune Lehmann , Benny E. Lautrup , Andrew D. Jackson
‹ Prev 1 3 4 5 6 7 10 Next ›