中文
相关论文

相关论文: Who and What Links to the Internet Archive

200 篇论文

Search engine is main access to the largest information source in this world, Internet. Now Internet is changing every aspect of our life. Information retrieval service may be its most important services. But for common user, internet…

网络与互联网体系结构 · 计算机科学 2007-05-23 Wang Liang , Guo Yi-Ping , Fang Ming

Verifiability is a core content policy of Wikipedia: claims that are likely to be challenged need to be backed by citations. There are millions of articles available online and thousands of new articles are released each month. For this…

Upon replay, JavaScript on archived web pages can generate recurring HTTP requests that lead to unnecessary traffic to the web archive. In one example, an archived page averaged more than 1000 requests per minute. These requests are not…

网络与互联网体系结构 · 计算机科学 2022-12-02 Kritika Garg , Himarsha R. Jayanetti , Sawood Alam , Michele C. Weigle , Michael L. Nelson

Despite the advance of the Open Access (OA) movement, most scholarly production can only be accessed through a paywall. We conduct an international survey among researchers (N=3,304) to measure the willingness and motivations to use (or not…

数字图书馆 · 计算机科学 2022-12-13 Francisco Segado-Boj , Juan Martin-Quevedo , Juan-Jose Prieto-Gutierrez

Web archives are large longitudinal collections that store webpages from the past, which might be missing on the current live Web. Consequently, temporal search over such collections is essential for finding prominent missing webpages and…

信息检索 · 计算机科学 2017-02-07 Helge Holzmann , Wolfgang Nejdl , Avishek Anand

Recently developed information communication technologies, particularly the Internet, have affected how we, both as individuals and as a society, create, store, and recall information. Internet also provides us with a great opportunity to…

物理与社会 · 物理学 2023-01-05 Ruth García-Gavilanes , Anders Mollgaard , Milena Tsvetkova , Taha Yasseri

Webpages change over time, and web archives hold copies of historical versions of webpages. Users of web archives, such as journalists, want to find and view changes on webpages over time. However, the current search interfaces for web…

信息检索 · 计算机科学 2023-05-02 Lesley Frew , Michael L. Nelson , Michele C. Weigle

Wikipedia is one of the most popular sites on the Web, with millions of users relying on it to satisfy a broad range of information needs every day. Although it is crucial to understand what exactly these needs are in order to be able to…

社会与信息网络 · 计算机科学 2017-03-17 Philipp Singer , Florian Lemmerich , Robert West , Leila Zia , Ellery Wulczyn , Markus Strohmaier , Jure Leskovec

Automated analysis of privacy policies has proved a fruitful research direction, with developments such as automated policy summarization, question answering systems, and compliance detection. Prior research has been limited to analysis of…

计算机与社会 · 计算机科学 2021-07-22 Ryan Amos , Gunes Acar , Eli Lucherini , Mihir Kshirsagar , Arvind Narayanan , Jonathan Mayer

AI-powered search systems are emerging as new information gatekeepers, fundamentally transforming how users access news and information. Despite their growing influence, the citation patterns of these systems remain poorly understood. We…

信息检索 · 计算机科学 2025-07-09 Kai-Cheng Yang

Backup or preservation of websites is often not considered until after a catastrophic event has occurred. In the face of complete website loss, "lazy" webmasters or concerned third parties may be able to recover some of their website from…

信息检索 · 计算机科学 2011-11-09 Frank McCown , Joan A. Smith , Michael L. Nelson , Johan Bollen

Memento aggregators enable users to query multiple web archives for captures of a URI in time through a single HTTP endpoint. While this one-to-many access point is useful for researchers and end-users, aggregators are in a position to…

数字图书馆 · 计算机科学 2023-01-10 Mat Kelly

A heightened interest in the presence of the past has given rise to the new field of memory studies, but there is a lack of search and research tools to support studying how and why the past is evoked in diachronic discourses. Searching for…

信息检索 · 计算机科学 2017-10-04 Alex Olieman , Kaspar Beelen , Jaap Kamps

As one of the Web's primary multilingual knowledge sources, Wikipedia is read by millions of people across the globe every day. Despite this global readership, little is known about why users read Wikipedia's various language editions. To…

计算机与社会 · 计算机科学 2018-12-04 Florian Lemmerich , Diego Sáez-Trumper , Robert West , Leila Zia

The Internet Yellow Pages (IYP) aggregates information from multiple sources about Internet routing into a unified, graph-based knowledge base. However, querying it requires knowledge of the Cypher language and the exact IYP schema, thus…

网络与互联网体系结构 · 计算机科学 2025-09-25 Vasilis Andritsoudis , Pavlos Sermpezis , Ilias Dimitriadis , Athena Vakali

Twitter is among the commonest sources of data employed in social media research mainly because of its convenient APIs to collect tweets. However, most researchers do not have access to the expensive Firehose and Twitter Historical Archive,…

计算机与社会 · 计算机科学 2016-11-28 Daniel Gayo-Avello

When you have a question, the most effective way to have the question answered is to directly connect with experts on the topic and have a conversation with them. Prior to the invention of writing, this was the only way. Although effective,…

信息检索 · 计算机科学 2024-12-30 Jimmy Lin , Pankaj Gupta , Will Horn , Gilad Mishne

Web is a primary and essential service to share information among users and organizations at present all over the world. Despite the current significance of such a kind of traffic on the Internet, the so-called Surface Web traffic has been…

The World Wide Web (WWW) allows the people to share the information (data) from the large database repositories globally. The amount of information grows billions of databases. We need to search the information will specialize tools known…

人工智能 · 计算机科学 2011-02-07 G. Madhu , Dr. A. Govardhan , Dr. T. V. Rajinikanth

As the amount of personal information stored at remote service providers increases, so does the danger of data theft. When connections to remote services are made in the clear and authenticated sessions are kept using HTTP cookies, data…

密码学与安全 · 计算机科学 2015-03-13 Claude Castelluccia , Emiliano De Cristofaro , Daniele Perito