English
Related papers

Related papers: Web2Wiki: Characterizing Wikipedia Linking Across …

200 papers

We propose a dynamic map of knowledge generated from Wikipedia pages and the Web URLs contained therein. GalaxySearch provides answers to the questions we don't know how to ask, by constructing a semantic network of the most relevant pages…

Social and Information Networks · Computer Science 2012-04-17 Hauke Fuehres , Peter A. Gloor , Michael Henninger , Reto Kleeb , Keiichi Nemoto

Wikipedia's perceived high quality and broad language coverage have established it as a fundamental resource in NLP. However, in recent years, such assumptions of high quality have become the subject of scrutiny in low-resource and…

This article analyzes one month of edits to Wikipedia in order to examine the role of users editing multiple language editions (referred to as multilingual users). Such multilingual users may serve an important function in diffusing…

Computers and Society · Computer Science 2014-05-13 Scott A. Hale

We use the directed networks between articles of 24 Wikipedia language editions for producing the Wikipedia Ranking of World Universities (WRWU) using PageRank, 2DRank and CheiRank algorithms. This approach allows to incorporate various…

Social and Information Networks · Computer Science 2016-03-16 José Lages , Antoine Patt , Dima L. Shepelyansky

Wikipedia is among the largest examples of collective intelligence on the Web with over 61 million articles covering over 320 languages. Although edited and maintained by an active workforce of human volunteers, Wikipedia is highly reliant…

Human-Computer Interaction · Computer Science 2025-09-29 Neal Reeves , Elena Simperl

Several hundred Wikipedia articles are deleted every day because they lack sufficient significance to be included in the encyclopedia. We collect a dataset of deleted articles and analyze them to determine whether or not the deletions were…

Computers and Society · Computer Science 2013-05-24 Bluma S. Gelley

Wikipedia is one of the largest online encyclopedias, which relies on scientific publications as authoritative sources. The increasing prevalence of open access (OA) publishing has expanded the public availability of scientific knowledge;…

Digital Libraries · Computer Science 2025-10-17 Puyu Yang , Vincent Traag , Rodrigo Costas , Giovanni Colavizza

Universities face increasing demands to improve their visibility, public outreach, and online presence. There is a broad consensus that scientific reputation significantly increases the attention universities receive. However, in most cases…

Digital Libraries · Computer Science 2023-07-12 Wenceslao Arroyo-Machado , Adrián A. Díaz-Faes , Enrique Herrera-Viedma , Rodrigo Costas

The Wikipedia category graph serves as the taxonomic backbone for large-scale knowledge graphs like YAGO or Probase, and has been used extensively for tasks like entity disambiguation or semantic similarity estimation. Wikipedia's…

Information Retrieval · Computer Science 2019-07-01 Nicolas Heist , Heiko Paulheim

Knowledge bases are prevalent in various domains and have been widely used in a large number of real applications such as applications in online encyclopedia, social media, biomedical fields, bibliographical networks. Due to their great…

Databases · Computer Science 2021-07-09 Feixiang Wang , Yixiang Fang , Yan Song , Shuang Li , Xinyun Chen

We perform a large-scale analysis of third-party trackers on the World Wide Web from more than 3.5 billion web pages of the CommonCrawl 2012 corpus. We extract a dataset containing more than 140 million third-party embeddings in over 41…

Social and Information Networks · Computer Science 2016-08-01 Sebastian Schelter , Jérôme Kunegis

The DBpedia project extracts structured information from Wikipedia and makes it available on the web. Information is gathered mainly with the help of infoboxes that contain structured information of the Wikipedia article. A lot of…

Information Retrieval · Computer Science 2012-05-21 Daniel Hienert , Francesco Luciano

This study concerned the active use of Wikipedia as a teaching tool in the classroom in higher education, trying to identify different usage profiles and their characterization. A questionnaire survey was administrated to all full-time and…

Social and Information Networks · Computer Science 2018-01-23 Julià Minguillón , Eduard Aibar , Maura Lerga , Josep Lladós , Antoni Meseguer-Artola

In this paper we present statistical analysis of English texts from Wikipedia. We try to address the issue of language complexity empirically by comparing the simple English Wikipedia (Simple) to comparable samples of the main English…

Computation and Language · Computer Science 2023-01-05 Taha Yasseri , András Kornai , János Kertész

Research on vandalism in Wikipedia has been of interest for the last decade. This paper performs a literature review on the subject, with the goal of identifying the main research topics and approaches, methods and techniques used. 67…

Digital Libraries · Computer Science 2016-06-20 Jesús Tramullas , Piedad Garrido-Picazo , Ana I. Sánchez-Casabón

Wikipedia -- like most peer production communities -- suffers from a basic problem: the amount of work that needs to be done (articles to be created and improved) exceeds the available resources (editor effort). Recommender systems have…

Computers and Society · Computer Science 2022-08-18 Mo Houtti , Isaac Johnson , Joel Cepeda , Soumya Khandelwal , Aviral Bhatnagar , Loren Terveen

Every day millions of people read Wikipedia. When navigating the vast space of available topics using hyperlinks, readers describe trajectories on the article network. Understanding these navigation patterns is crucial to better serve…

Computers and Society · Computer Science 2022-01-10 Akhil Arora , Martin Gerlach , Tiziano Piccardi , Alberto García-Durán , Robert West

Links are a fundamental part of information networks, turning isolated pieces of knowledge into a network of information that is much richer than the sum of its parts. However, adding a new link to the network is not trivial: it requires…

Computation and Language · Computer Science 2024-10-08 Tomás Feith , Akhil Arora , Martin Gerlach , Debjit Paul , Robert West

Finding relevant information from large document collections such as the World Wide Web is a common task in our daily lives. Estimation of a user's interest or search intention is necessary to recommend and retrieve relevant information…

Information Retrieval · Computer Science 2016-12-09 Manuel J. A. Eugster , Tuukka Ruotsalo , Michiel M. Spapé , Oswald Barral , Niklas Ravaja , Giulio Jacucci , Samuel Kaski

The Library of Babel, described by Jorge Luis Borges, stores an enormous amount of information. The Library exists {\it ab aeterno}. Wikipedia, a free online encyclopaedia, becomes a modern analogue of such a Library. Information retrieval…

Information Retrieval · Computer Science 2010-11-15 A. O. Zhirov , O. V. Zhirov , D. L. Shepelyansky