English
Related papers

Related papers: Annotations for Exploring Food Tweets From Multipl…

200 papers

Analysis of short text, such as social media posts, is extremely difficult because of their inherent brevity. In addition to classifying topics of such posts, a common downstream task is grouping the authors of these documents for…

Information Retrieval · Computer Science 2022-06-20 Graham Tierney , Christopher Bail , Alexander Volfovsky

Large language models (LLMs) offer new opportunities for scalable analysis of online discourse. Yet their use in multilingual social science research remains constrained by model size, cost and linguistic bias. We develop a lightweight,…

Computation and Language · Computer Science 2025-12-30 Andrea Nasuto , Stefano Maria Iacus , Francisco Rowe , Devika Jain

Social scientists increasingly use demographically stratified social media data to study the attitudes, beliefs, and behavior of the general public. To facilitate such analyses, we construct, validate, and release publicly the…

Computation and Language · Computer Science 2024-03-12 Lorenzo Lupo , Paul Bose , Mahyar Habibi , Dirk Hovy , Carlo Schwarz

We present MM-Food-100K, a public 100,000-sample multimodal food intelligence dataset with verifiable provenance. It is a curated approximately 10% open subset of an original 1.2 million, quality-accepted corpus of food images annotated for…

Artificial Intelligence · Computer Science 2025-08-15 Yi Dong , Yusuke Muraoka , Scott Shi , Yi Zhang

This paper introduces a study on tweet sentiment classification. Our task is to classify a tweet as either positive or negative. We approach the problem in two steps, namely embedding and classifying. Our baseline methods include several…

Computation and Language · Computer Science 2021-10-01 Tommaso Macrì , Freya Murphy , Yunfan Zou , Yves Zumbach

Stance classification aims to identify, for a particular issue under discussion, whether the speaker or author of a conversational turn has Pro (Favor) or Con (Against) stance on the issue. Detecting stance in tweets is a new task proposed…

Computation and Language · Computer Science 2018-01-29 Amita Misra , Brian Ecker , Theodore Handleman , Nicolas Hahn , Marilyn Walker

Social Internet content plays an increasingly critical role in many domains, including public health, disaster management, and politics. However, its utility is limited by missing geographic information; for example, fewer than 1.6% of…

Social and Information Networks · Computer Science 2013-11-19 Reid Priedhorsky , Aron Culotta , Sara Y. Del Valle

We propose to analyze the level of recommendation and spreading in the sharing of scientific papers on Twitter to understand the interactions of communities around papers and to develop the "Community of Attention Network" (CAN). In this…

Digital Libraries · Computer Science 2020-06-16 Ronaldo Ferreira Araujo

The wide use of social media and digital technologies facilitates sharing various news and information about events and activities. Despite sharing positive information misleading and false information is also spreading on social media.…

Computation and Language · Computer Science 2022-07-18 Prerona Tarannum , Firoj Alam , Md. Arid Hasan , Sheak Rashed Haider Noori

Social media plays a significant role in cross-cultural communication. A vast amount of this occurs in code-mixed and multilingual form, posing a significant challenge to Natural Language Processing (NLP) tools for processing such…

Computation and Language · Computer Science 2026-01-21 Dwip Dalal , Vivek Srivastava , Mayank Singh

Customer-provided reviews have become an important source of information for business owners and other customers alike. However, effectively analyzing millions of unstructured reviews remains challenging. While large language models (LLMs)…

Computation and Language · Computer Science 2026-02-25 Vishal Patil , Shree Vaishnavi Bacha , Revanth Yamani , Yidan Sun , Mayank Kejriwal

With the exponential growth in the usage of social media to share live updates about life, taking pictures has become an unavoidable phenomenon. Individuals unknowingly create a unique knowledge base with these images. The food images, in…

Computer Vision and Pattern Recognition · Computer Science 2020-03-20 Nitish Nag , Bindu Rajanna , Ramesh Jain

Since state-of-the-art approaches to offensive language detection rely on supervised learning, it is crucial to quickly adapt them to the continuously evolving scenario of social media. While several approaches have been proposed to tackle…

Computation and Language · Computer Science 2022-10-17 Elisa Leonardelli , Stefano Menini , Alessio Palmero Aprosio , Marco Guerini , Sara Tonelli

The trends and reactions of the general public towards global events can be analyzed using data from social platforms, including Twitter. The number of tweets has been reported to help detect variations in communication traffic within…

Applications · Statistics 2024-06-05 Sneha Jha , Dharmendra Saraswat , Mark D. Ward

Social Media users tend to mention entities when reacting to news events. The main purpose of this work is to create entity-centric aggregations of tweets on a daily basis. By applying topic modeling and sentiment analysis, we create data…

Social and Information Networks · Computer Science 2018-01-25 João Oliveira , Mike Pinto , Pedro Saleiro , Jorge Teixeira

In this paper, we, as the DS@GT team for CLEF 2025 CheckThat! Task 4a Scientific Web Discourse Detection, present the methods we explored for this task. For this multiclass classification task, we determined if a tweet contained a…

Computation and Language · Computer Science 2025-07-09 Ayush Parikh , Hoang Thanh Thanh Truong , Jeanette Schofield , Maximilian Heil

Turkish Wikipedia Named-Entity Recognition and Text Categorization (TWNERTC) dataset is a collection of automatically categorized and annotated sentences obtained from Wikipedia. We constructed large-scale gazetteers by using a graph…

Computation and Language · Computer Science 2017-02-10 H. Bahadir Sahin , Caglar Tirkaz , Eray Yildiz , Mustafa Tolga Eren , Ozan Sonmez

Free-text responses are commonly collected in psychological studies, providing rich qualitative insights that quantitative measures may not capture. Labeling curated topics of research interest in free-text data by multiple trained human…

Nowadays social media platforms such as Twitter provide a great opportunity to understand public opinion of climate change compared to traditional survey methods. In this paper, we constructed a massive climate change Twitter dataset and…

Computation and Language · Computer Science 2021-12-01 Zhongkai Shangguan , Zihe Zheng , Lei Lin

In this work, we present a Web-based annotation tool `Relation Triplets Extractor' \footnote{https://abera87.github.io/annotate/} (RTE) for annotating relation triplets from the text. Relation extraction is an important task for extracting…

Computation and Language · Computer Science 2021-08-19 Ankan Mullick , Animesh Bera , Tapas Nayak