中文
相关论文

相关论文: Text-Based Ideal Points

200 篇论文

We propose a novel training and inference method for detecting political bias in long text content such as newspaper opinion articles. Obtaining long text data and annotations at sufficient scale for training is difficult, but it is…

计算与语言 · 计算机科学 2019-11-20 Aditya Saligrama

Ideological leanings of an individual can often be gauged by the sentiment one expresses about different issues. We propose a simple framework that represents a political ideology as a distribution of sentiment polarities towards a set of…

计算与语言 · 计算机科学 2018-10-31 Sumit Bhatia , Deepak P

We propose a novel supervised learning approach for political ideology prediction (PIP) that is capable of predicting out-of-distribution inputs. This problem is motivated by the fact that manual data-labeling is expensive, while…

机器学习 · 计算机科学 2023-02-02 Chen Chen , Dylan Walker , Venkatesh Saligrama

Current cross-platform social media analyses primarily focus on the textual features of posts, often lacking multimodal analysis due to past technical limitations. This study addresses this gap by examining how U.S. legislators in the 118th…

计算机与社会 · 计算机科学 2025-09-17 Weihong Qi , Anushka Dave , Chen Ling

The election forecasting 'industry' is a growing one, both in the volume of scholars producing forecasts and methodological diversity. In recent years a new approach has emerged that relies on social media and particularly Twitter data to…

计算机与社会 · 计算机科学 2015-05-08 Pete Burnap , Rachel Gibson , Luke Sloan , Rosalynd Southern , Matthew Williams

The increasing use of social networks generates enormous amounts of data that can be used for many types of analysis. Some of these data have temporal and geographical information, which can be used for comprehensive examination. In this…

社会与信息网络 · 计算机科学 2012-10-16 Augusto Dias Pereira dos Santos , Leandro Krug Wives , Luis Otavio Alvares

Social media serves as a critical medium in modern politics because it both reflects politicians' ideologies and facilitates communication with younger generations. We present MultiParTweet, a multilingual tweet corpus from X that connects…

计算与语言 · 计算机科学 2025-12-15 Mevlüt Bagci , Ali Abusaleh , Daniel Baumartz , Giueseppe Abrami , Maxim Konca , Alexander Mehler

The quantitative analysis of political ideological positions is a difficult task. In the past, various literature focused on parliamentary voting data of politicians, party manifestos and parliamentary speech to estimate political…

计算与语言 · 计算机科学 2024-05-14 Ken Kato , Annabelle Purnomo , Christopher Cochrane , Raeid Saqur

Growing literature has shown that NLP systems may encode social biases; however, the political bias of summarization models remains relatively unknown. In this work, we use an entity replacement method to investigate the portrayal of…

计算与语言 · 计算机科学 2023-10-23 Karen Zhou , Chenhao Tan

Statistical topic models provide a general data-driven framework for automated discovery of high-level knowledge from large collections of text documents. While topic models can potentially discover a broad range of themes in a data set,…

人工智能 · 计算机科学 2008-08-08 Chaitanya Chemudugunta , Padhraic Smyth , Mark Steyvers

In recent days, the amount of Cyber Security text data shared via social media resources mainly Twitter has increased. An accurate analysis of this data can help to develop cyber threat situational awareness framework for a cyber threat.…

计算与语言 · 计算机科学 2020-04-02 Simran K , Prathiksha Balakrishna , Vinayakumar R , Soman KP

The authorship attribution is a problem of considerable practical and technical interest. Several methods have been designed to infer the authorship of disputed documents in multiple contexts. While traditional statistical methods based…

计算与语言 · 计算机科学 2018-03-28 Jeaneth Machicao , Edilson A. Corrêa , Gisele H. B. Miranda , Diego R. Amancio , Odemir M. Bruno

The increasing growth of social media provides us with an instant opportunity to be informed of the opinions of a large number of politically active individuals in real-time. We can get an overall idea of the ideologies of these individuals…

人机交互 · 计算机科学 2024-11-08 Sultan Ahmed , Salman Rakin , Khadija Urmi , Chandan Kumar Nag , Md. Mostofa Akbar

We use instruction-tuned Large Language Models (LLMs) like GPT-4, Llama 3, MiXtral, or Aya to position political texts within policy and ideological spaces. We ask an LLM where a tweet or a sentence of a political text stands on the focal…

计算与语言 · 计算机科学 2024-09-06 Gaël Le Mens , Aina Gallego

In this manuscript, we analyze the interaction network on Twitter among members of the 117th U.S. Congress to assess the visibility of political leaders and explore how systemic properties and node attributes influence the formation of…

应用统计 · 统计学 2024-09-17 Carolina Luque , Juan Sosa

The increasing popularity of Twitter renders improved trustworthiness and relevance assessment of tweets much more important for search. However, given the limitations on the size of tweets, it is hard to extract measures for ranking from…

信息检索 · 计算机科学 2013-08-13 Srijith Ravikumar , Kartik Talamadupula , Raju Balakrishnan , Subbarao Kambhampati

Social media has become an emerging alternative to opinion polls for public opinion collection, while it is still posing many challenges as a passive data source, such as structurelessness, quantifiability, and representativeness. Social…

社会与信息网络 · 计算机科学 2020-05-26 Zhaoya Gong , Tengteng Cai , Jean-Claude Thill , Scott Hale , Mark Graham

Large-scale data from social media have a significant potential to describe complex phenomena in real world and to anticipate collective behaviors such as information spreading and social trends. One specific case of study is represented by…

物理与社会 · 物理学 2015-07-15 Young-Ho Eom , Michelangelo Puliga , Jasmina Smailović , Igor Mozetič , Guido Caldarelli

Topic modelling is a text mining technique for identifying salient themes from a number of documents. The output is commonly a set of topics consisting of isolated tokens that often co-occur in such documents. Manual effort is often…

计算与语言 · 计算机科学 2024-04-26 Lowri Williams , Eirini Anthi , Laura Arman , Pete Burnap

A word embedding is a low-dimensional, dense and real- valued vector representation of a word. Word embeddings have been used in many NLP tasks. They are usually gener- ated from a large text corpus. The embedding of a word cap- tures both…

计算与语言 · 计算机科学 2017-08-15 Quanzhi Li , Sameena Shah , Xiaomo Liu , Armineh Nourbakhsh