中文
相关论文

相关论文: Intriguing behavior when testing the impact of quo…

200 篇论文

Search-engine date filters are widely used to enforce pre-cutoff retrieval in retrospective evaluations of search-augmented forecasters. We show this approach is unreliable across two major search engines: auditing Google Search's before:…

计算与语言 · 计算机科学 2026-04-22 Ali El Lahib , Ying-Jieh Xia , Zehan Li , Yuxuan Wang , Xinyu Pi

We applied a set of standard bibliometric indicators to monitor the scientific state-of-arte of 500 universities worldwide and constructed a ranking on the basis of these indicators (Leiden Ranking 2010). We find a dramatic and hitherto…

数字图书馆 · 计算机科学 2010-12-24 Anthony F. J. van Raan , Thed N. van Leeuwen , Martijn S. Visser

The intriguing law of anomalous numbers, also named Benford's law, states that the significant digits of data follow a logarithmic distribution favoring the smallest values. In this work, we test the compliance with this law of the atomic…

原子物理 · 物理学 2024-05-30 Jean-Christophe Pain , Yuri Ralchenko

Our goal is to distinguish between the following two hypotheses: (A) The Internet will remain disproportionately in English and will, over time, cause more people to learn English as second language and thus solidify the role of English as…

计算机与社会 · 计算机科学 2007-05-23 Neil Gandal , Carl Shapiro

Benford's law is widely used for fraud-detection nowadays. The underlying assumption for using the law is that a "regular" dataset follows the significant digit phenomenon. In this paper, we address the scenario where a shrewd fraudster…

应用统计 · 统计学 2021-05-21 Javad Kazemitabar

The Newcomb-Benford law, also known as the first-digit law, gives the probability distribution associated with the first digit of a dataset, so that, for example, the first significant digit has a probability of $30.1$ % of being $1$ and…

科普物理 · 物理学 2021-08-25 Andrea Burgos , Andrés Santos

In this work, we introduce a novel metric for auditing group fairness in ranked lists. Our approach offers two benefits compared to the state of the art. First, we offer a blueprint for modeling of user attention. Rather than assuming a…

计算机与社会 · 计算机科学 2019-05-14 Piotr Sapiezynski , Wesley Zeng , Ronald E. Robertson , Alan Mislove , Christo Wilson

The correlation between the demographics of users and the text they write has been investigated through literary texts and, more recently, social media. However, differences pertaining to language use in search engines has not been…

计算机与社会 · 计算机科学 2018-05-24 Elad Yom-Tov

Search engines decide what we see for a given search query. Since many people are exposed to information through search engines, it is fair to expect that search engines are neutral. However, search engine results do not necessarily cover…

信息检索 · 计算机科学 2023-02-07 Gizem Gezici , Aldo Lipani , Yucel Saygin , Emine Yilmaz

The proliferation of surveys and review articles in academic journals has impacted citation metrics like impact factor and h-index, skewing evaluations of journal and researcher quality. This work investigates the implications of this…

数字图书馆 · 计算机科学 2025-04-09 Jesus S. Aguilar-Ruiz

It is tempting to treat frequency trends from the Google Books data sets as indicators of the "true" popularity of various words and phrases. Doing so allows us to draw quantitatively strong conclusions about the evolution of cultural…

物理与社会 · 物理学 2020-05-28 Eitan Adam Pechenick , Christopher M. Danforth , Peter Sheridan Dodds

Nonextensive statistics, characterized by a nonextensive parameter $q$, is a promising and practically useful generalization of the Boltzmann statistics to describe power-law behaviors from physical and social observations. We here explore…

数据分析、统计与概率 · 物理学 2011-03-07 Lijing Shao , Bo-Qiang Ma

We carried out a retrieval effectiveness test on the three major web search engines (i.e., Google, Microsoft and Yahoo). In addition to relevance judgments, we classified the results according to their commercial intent and whether or not…

信息检索 · 计算机科学 2015-11-19 Dirk Lewandowski

In this paper, we propose a measure to assess scientific impact that discounts self-citations and does not require any prior knowledge on the their distribution among publications. This index can be applied to both researchers and journals.…

数字图书馆 · 计算机科学 2013-10-17 Emilio Ferrara , Alfonso E. Romero

When interacting with information retrieval (IR) systems, users, affected by confirmation biases, tend to select search results that confirm their existing beliefs on socially significant contentious issues. To understand the judgments and…

信息检索 · 计算机科学 2024-06-18 Ben Wang , Jiqun Liu

The impact of scientific publications has traditionally been expressed in terms of citation counts. However, scientific activity has moved online over the past decade. To better capture scientific impact in the digital era, a variety of new…

数字图书馆 · 计算机科学 2009-06-30 Johan Bollen , Herbert Van de Sompel , Aric Hagberg , Ryan Chute

There is inherent information captured in the order in which we write words in a list. The orderings of binomials --- lists of two words separated by `and' or `or' --- has been studied for more than a century. These binomials are common…

社会与信息网络 · 计算机科学 2020-03-10 Katherine Van Koevering , Austin R. Benson , Jon Kleinberg

Purpose: To compare five major Web search engines (Google, Yahoo, MSN, Ask.com, and Seekport) for their retrieval effectiveness, taking into account not only the results but also the results descriptions. Design/Methodology/Approach: The…

信息检索 · 计算机科学 2015-11-19 Dirk Lewandowski

This paper extends Becker (1957)'s outcome test of discrimination to settings where a (human or algorithmic) decision-maker produces a ranked list of candidates. Ranked lists are particularly relevant in the context of online platforms that…

计量经济学 · 经济学 2021-11-16 Jonathan Roth , Guillaume Saint-Jacques , YinYin Yu

The main objective of this paper is to empirically test whether the identification of highly-cited documents through Google Scholar is feasible and reliable. To this end, we carried out a longitudinal analysis (1950 to 2013), running a…