中文
相关论文

相关论文: TweeNLP: A Twitter Exploration Portal for Natural …

200 篇论文

Whether it is in the form of transcribed conversations, blog posts, or tweets, qualitative data provides a reader with rich insight into both the overarching trends as well as the diversity of human ideas expressed through text. Handling…

人机交互 · 计算机科学 2022-09-27 Huyen N. Nguyen , Tommy Dang , Kathleen A. Bowe

With the rise in popularity of public social media and micro-blogging services, most notably Twitter, the people have found a venue to hear and be heard by their peers without an intermediary. As a consequence, and aided by the public…

计算与语言 · 计算机科学 2016-06-21 Prashanth Vijayaraghavan , Soroush Vosoughi , Deb Roy

Recent advances in NLP have improved our ability to understand the nuanced worldviews of online communities. Existing research focused on probing ideological stances treats liberals and conservatives as separate groups. However, this fails…

计算与语言 · 计算机科学 2024-02-05 Zihao He , Ashwin Rao , Siyi Guo , Negar Mokhberian , Kristina Lerman

Text from social media provides a set of challenges that can cause traditional NLP approaches to fail. Informal language, spelling errors, abbreviations, and special characters are all commonplace in these posts, leading to a prohibitively…

机器学习 · 计算机科学 2016-05-18 Bhuwan Dhingra , Zhong Zhou , Dylan Fitzpatrick , Michael Muehl , William W. Cohen

In this paper, we present an experiment on using deep learning and transfer learning techniques for emotion analysis in tweets and suggest a method to interpret our deep learning models. The proposed approach for emotion analysis combines a…

计算与语言 · 计算机科学 2020-12-14 Yasas Senarath , Uthayasanker Thayasivam

In this paper, we describe the Lithium Natural Language Processing (NLP) system - a resource-constrained, high- throughput and language-agnostic system for information extraction from noisy user generated text on social media. Lithium NLP…

人工智能 · 计算机科学 2017-07-14 Preeti Bhargava , Nemanja Spasojevic , Guoning Hu

This research is aimed to solve the tweet/user geolocation prediction task and provide a flexible methodology for the geotagging of textual big data. The suggested approach implements neural networks for natural language processing (NLP) to…

计算与语言 · 计算机科学 2025-01-13 Kateryna Lutsai , Christoph H. Lampert

During natural or man-made disasters, humanitarian response organizations look for useful information to support their decision-making processes. Social media platforms such as Twitter have been considered as a vital source of useful…

计算与语言 · 计算机科学 2016-10-06 Dat Tien Nguyen , Shafiq Joty , Muhammad Imran , Hassan Sajjad , Prasenjit Mitra

Social media user profiling through content analysis is crucial for tasks like misinformation detection, engagement prediction, hate speech monitoring, and user behavior modeling. However, existing profiling techniques, including tweet…

社会与信息网络 · 计算机科学 2025-05-12 Vahid Rahimzadeh , Ali Hamzehpour , Azadeh Shakery , Masoud Asadpour

As one of the most extensive social networking services, Twitter has more than 300 million active users as of 2022. Among its many functions, Twitter is now one of the go-to platforms for consumers to share their opinions about products or…

计算与语言 · 计算机科学 2022-09-30 Shengyang Wu , Yi Gao

Despite its importance, the time variable has been largely neglected in the NLP and language model literature. In this paper, we present TimeLMs, a set of language models specialized on diachronic Twitter data. We show that a continual…

计算与语言 · 计算机科学 2022-04-04 Daniel Loureiro , Francesco Barbieri , Leonardo Neves , Luis Espinosa Anke , Jose Camacho-Collados

Large-scale data sets on scholarly publications are the basis for a variety of bibliometric analyses and natural language processing (NLP) applications. Especially data sets derived from publication's full-text have recently gained…

数字图书馆 · 计算机科学 2023-11-06 Tarek Saier , Johan Krause , Michael Färber

Natural Language Processing offers new insights into language data across almost all disciplines and domains, and allows us to corroborate and/or challenge existing knowledge. The primary hurdles to widening participation in and use of…

计算与语言 · 计算机科学 2021-05-31 Rebekah Baglini , Arthur Hjorth

User simulators are often used to generate large amounts of data for various tasks such as generation, training, and evaluation. However, existing approaches concentrate on collective behaviors or interactive systems, struggling with tasks…

信息检索 · 计算机科学 2026-02-27 Bingrui Jin , Kunyao Lan , Mengyue Wu

Microblogging platforms, of which Twitter is a representative example, are valuable information sources for market screening and financial models. In them, users voluntarily provide relevant information, including educated knowledge on…

Social media platforms contain a great wealth of information which provides opportunities for us to explore hidden patterns or unknown correlations, and understand people's satisfaction with what they are discussing. As one showcase, in…

信息检索 · 计算机科学 2017-05-24 Zhengkui Wang , Guangdong Bai , Soumyadeb Chowdhury , Quanqing Xu , Zhi Lin Seow

Nowadays social media platforms such as Twitter provide a great opportunity to understand public opinion of climate change compared to traditional survey methods. In this paper, we constructed a massive climate change Twitter dataset and…

计算与语言 · 计算机科学 2021-12-01 Zhongkai Shangguan , Zihe Zheng , Lei Lin

Cluster analysis is a field of data analysis that extracts underlying patterns in data. One application of cluster analysis is in text-mining, the analysis of large collections of text to find similarities between documents. We used a…

机器学习 · 统计学 2014-08-26 Daniel Godfrey , Caley Johns , Carl Meyer , Shaina Race , Carol Sadek

Breast cancer is a significant public health concern and is the leading cause of cancer-related deaths among women. Despite advances in breast cancer treatments, medication non-adherence remains a major problem. As electronic health records…

计算与语言 · 计算机科学 2024-07-30 Seibi Kobara , Alireza Rafiei , Masoud Nateghi , Selen Bozkurt , Rishikesan Kamaleswaran , Abeed Sarker

A word embedding is a low-dimensional, dense and real- valued vector representation of a word. Word embeddings have been used in many NLP tasks. They are usually gener- ated from a large text corpus. The embedding of a word cap- tures both…

计算与语言 · 计算机科学 2017-08-15 Quanzhi Li , Sameena Shah , Xiaomo Liu , Armineh Nourbakhsh