English
Related papers

Related papers: Preserving the Ephemeral: Instagram Story Archivin…

200 papers

To improve the reading experience, many news sites organize news into topical collections, called stories. In this work, we present an approach for implementing real-time story identification for a news monitoring system that automatically…

Computation and Language · Computer Science 2025-08-13 Tadej Škvorc , Nikola Ivačič , Sebastjan Hribar , Marko Robnik-Šikonja

The spread of false rumours during emergencies can jeopardise the well-being of citizens as they are monitoring the stream of news from social media to stay abreast of the latest updates. In this paper, we describe the methodology we have…

Social and Information Networks · Computer Science 2015-04-21 Arkaitz Zubiaga , Maria Liakata , Rob Procter , Kalina Bontcheva , Peter Tolmie

The global popularity of microblogs has led to an increasing accumulation of large volumes of text data on microblogging platforms such as Twitter. These corpora are untapped resources to understand social expressions on diverse subjects.…

Information Retrieval · Computer Science 2018-11-15 Tharindu Rukshan Bandaragoda , Daswin De Silva , Damminda Alahakoon

This work studies how to transform an album to vivid and coherent stories, a task we refer to as "album storytelling". While this task can help preserve memories and facilitate experience sharing, it remains an underexplored area in current…

Computer Vision and Pattern Recognition · Computer Science 2023-05-25 Munan Ning , Yujia Xie , Dongdong Chen , Zeyin Song , Lu Yuan , Yonghong Tian , Qixiang Ye , Li Yuan

Performance of neural models for named entity recognition degrades over time, becoming stale. This degradation is due to temporal drift, the change in our target variables' statistical properties over time. This issue is especially…

Computation and Language · Computer Science 2021-04-21 Shuguang Chen , Leonardo Neves , Thamar Solorio

Twitter is often the most up-to-date source for finding and tracking breaking news stories. Therefore, there is considerable interest in developing filters for tweet streams in order to track and summarize stories. This is a non-trivial…

Information Retrieval · Computer Science 2014-12-01 Igor Brigadir , Derek Greene , Pádraig Cunningham

Streams of user-generated content in social media exhibit patterns of collective attention across diverse topics, with temporal structures determined both by exogenous factors and endogenous factors. Teasing apart different topics and…

Physics and Society · Physics 2014-03-07 A. Panisson , L. Gauvin , M. Quaggiotto , C. Cattuto

Providing personalized recommendations in an environment where items exhibit ephemerality and temporal relevancy (e.g. in social media) presents a few unique challenges: (1) inductively understanding ephemeral appeal for items in a setting…

Social and Information Networks · Computer Science 2022-10-31 Frank Portman , Stephen Ragain , Ahmed El-Kishky

We uncover a previously unknown, ongoing astroturfing attack on the popularity mechanisms of social media platforms: ephemeral astroturfing attacks. In this attack, a chosen keyword or topic is artificially promoted by coordinated and…

Cryptography and Security · Computer Science 2021-03-15 Tuğrulcan Elmas , Rebekah Overdorf , Ahmed Furkan Özkalay , Karl Aberer

With the advent of social media, our online feeds increasingly consist of short, informal, and unstructured text. This textual data can be analyzed for the purpose of improving user recommendations and detecting trends. Instagram is one of…

Computation and Language · Computer Science 2019-09-25 Kim Hammar , Shatha Jaradat , Nima Dokoohaki , Mihhail Matskin

The problem of event extraction is a relatively difficult task for low resource languages due to the non-availability of sufficient annotated data. Moreover, the task becomes complex for tail (rarely occurring) labels wherein extremely less…

Information Retrieval · Computer Science 2020-11-23 Ayush Maheshwari , Hrishikesh Patel , Nandan Rathod , Ritesh Kumar , Ganesh Ramakrishnan , Pushpak Bhattacharyya

Web archiving is the process of collecting portions of the Web to ensure that the information is preserved for future exploitation. However, despite the increasing number of web archives worldwide, the absence of efficient and meaningful…

Digital Libraries · Computer Science 2018-10-25 Pavlos Fafalios , Helge Holzmann , Vaibhav Kasturia , Wolfgang Nejdl

We consider a task of scheduling a crawler to retrieve content from several sites with ephemeral content. A user typically loses interest in ephemeral content, like news or posts at social network groups, after several days or hours. Thus,…

Information Retrieval · Computer Science 2015-03-31 Konstantin Avrachenkov , Vivek Borkar

Work on social media rumour verification utilises signals from posts, their propagation and users involved. Other lines of work target identifying and fact-checking claims based on information from Wikipedia, or trustworthy news articles…

Computation and Language · Computer Science 2022-07-29 John Dougrez-Lewis , Elena Kochkina , M. Arana-Catania , Maria Liakata , Yulan He

People spend a significant amount of time trying to make sense of the internet, collecting content from a variety of sources and organizing it to make decisions and achieve their goals. While humans are able to fluidly iterate on collecting…

Human-Computer Interaction · Computer Science 2022-09-01 Andrew Kuznetsov , Joseph Chee Chang , Nathan Hahn , Napol Rachatasumrit , Bradley Breneisen , Julina Coupland , Aniket Kittur

We present a framework SCStory for online story discovery, that helps people digest rapidly published news article streams in real-time without human annotations. To organize news article streams into stories, existing approaches directly…

Computation and Language · Computer Science 2023-12-08 Susik Yoon , Yu Meng , Dongha Lee , Jiawei Han

In real-time, social media data strongly imprints world events, popular culture, and day-to-day conversations by millions of ordinary people at a scale that is scarcely conventionalized and recorded. Vitally, and absent from many standard…

This paper introduces a large collection of time series data derived from Twitter, postprocessed using word embedding techniques, as well as specialized fine-tuned language models. This data comprises the past five years and captures…

Computation and Language · Computer Science 2023-08-07 Daniel Loureiro , Kiamehr Rezaee , Talayeh Riahi , Francesco Barbieri , Leonardo Neves , Luis Espinosa Anke , Jose Camacho-Collados

User-generated social media data is constantly changing as new trends influence online discussion and personal information is deleted due to privacy concerns. However, most current NLP models are static and rely on fixed training data,…

Computation and Language · Computer Science 2023-05-17 Fatemehsadat Mireshghallah , Nikolai Vogler , Junxian He , Omar Florez , Ahmed El-Kishky , Taylor Berg-Kirkpatrick

To prevent the spread of disinformation on Instagram, we need to study the accounts and content of disinformation actors. However, due to their malicious nature, Instagram often bans accounts that are responsible for spreading…

Digital Libraries · Computer Science 2024-01-05 Rachel Zheng , Michele C. Weigle