English
Related papers

Related papers: Stereotypical gender actions can be extracted from…

200 papers

User-generated content (UGC) is characterised by frequent use of non-standard language, from spelling errors to expressive choices such as slang, character repetitions, and emojis. This makes evaluating UGC translation challenging: what…

Computation and Language · Computer Science 2026-05-13 Lydia Nishimwe , Benoît Sagot , Rachel Bawden

Semantic sentence embeddings are usually supervisedly built minimizing distances between pairs of embeddings of sentences labelled as semantically similar by annotators. Since big labelled datasets are rare, in particular for non-English…

Computation and Language · Computer Science 2021-10-06 Marco Di Giovanni , Marco Brambilla

This paper is focused on the computational analysis of collective discourse, a collective behavior seen in non-expert content contributions in online social media. We collect and analyze a wide range of real-world collective discourse…

Social and Information Networks · Computer Science 2012-04-18 Vahed Qazvinian , Dragomir R. Radev

This research draws upon cognitive psychology and information systems studies to anticipate user engagement and decision-making on digital platforms. By employing natural language processing (NLP) techniques and insights from cognitive bias…

Human-Computer Interaction · Computer Science 2023-07-28 Nimrod Dvir , Elaine Friedman , Suraj Commuri , Fan Yang , Jennifer Romano

We use structural topic modeling to examine racial bias in data collected to train models to detect hate speech and abusive language in social media posts. We augment the abusive language dataset by adding an additional feature indicating…

Computation and Language · Computer Science 2020-05-28 Thomas Davidson , Debasmita Bhattacharya

The goal of our research is to contribute information about how useful the crowd is at anticipating stereotypes that may be biasing a data set without a researcher's knowledge. The results of the crowd's prediction can potentially be used…

Human-Computer Interaction · Computer Science 2018-01-11 Zeyuan Hu , Julia Strout

With increasing globalization and immigration, various studies have estimated that about half of the world population is bilingual. Consequently, individuals concurrently use two or more languages or dialects in casual conversational…

Computation and Language · Computer Science 2022-11-01 Saurav K. Aryal , Howard Prioleau , Gloria Washington

Studies have shown that some Natural Language Processing (NLP) systems encode and replicate harmful biases with potential adverse ethical effects in our society. In this article, we propose an approach for identifying gender and racial…

Computation and Language · Computer Science 2022-04-13 Sean Matthews , John Hudzina , Dawn Sepehr

The movie recommender system typically leverages user feedback to provide personalized recommendations that align with user preferences and increase business revenue. This study investigates the impact of gender stereotypes on such systems…

Information Retrieval · Computer Science 2025-01-09 Falguni Roy , Yiduo Shen , Na Zhao , Xiaofeng Ding , Md. Omar Faruk

Currently, there is a surge of interest in fair Artificial Intelligence (AI) and Machine Learning (ML) research which aims to mitigate discriminatory bias in AI algorithms, e.g. along lines of gender, age, and race. While most research in…

Computers and Society · Computer Science 2021-07-29 Clarice Wang , Kathryn Wang , Andrew Bian , Rashidul Islam , Kamrun Naher Keya , James Foulds , Shimei Pan

Text is a vehicle to convey information that reflects the writer's linguistic style and communicative patterns. By studying these attributes, we can discover latent insights about the author and their underlying message. This article uses…

Computers and Society · Computer Science 2024-12-19 Deborah Gerhardt , Miriam Marcowitz-Bitton , W. Michael Schuster , Avshalom Elmalech , Omri Suissa , Moshe Mash

We propose Coactive Learning as a model of interaction between a learning system and a human user, where both have the common goal of providing results of maximum utility to the user. At each step, the system (e.g. search engine) receives a…

Machine Learning · Computer Science 2015-03-20 Pannaga Shivaswamy , Thorsten Joachims

This study delves into gender classification systems, shedding light on the interaction between social stereotypes and algorithmic determinations. Drawing on the "averageness theory," which suggests a relationship between a face's…

Computers and Society · Computer Science 2024-11-14 Miriam Doh , Anastasia Karagianni

Here we examine whether the personality dimension of openness to experience can be predicted from the individual google search history. By web scraping, individual text corpora (ICs) were generated from 214 participants with a mean number…

Computation and Language · Computer Science 2024-04-02 Markus J. Hofmann , Markus T. Jansen , Christoph Wigbels , Benny Briesemeister , Arthur M. Jacobs

Humans carry stereotypic tacit assumptions (STAs) (Prince, 1978), or propositional beliefs about generic concepts. Such associations are crucial for understanding natural language. We construct a diagnostic set of word prediction prompts to…

Computation and Language · Computer Science 2020-06-17 Nathaniel Weir , Adam Poliak , Benjamin Van Durme

Semantic representations are integral to natural language processing, psycholinguistics, and artificial intelligence. Although often derived from internet text, recent years have seen a rise in the popularity of behavior-based (e.g., free…

Computation and Language · Computer Science 2024-12-09 Zak Hussain , Rui Mata , Ben R. Newell , Dirk U. Wulff

Generating models from large data sets -- and determining which subsets of data to mine -- is becoming increasingly automated. However choosing what data to collect in the first place requires human intuition or experience, usually supplied…

Computers and Society · Computer Science 2014-05-20 Josh C. Bongard , Paul D. H. Hines , Dylan Conger , Peter Hurd , Zhenyu Lu

This paper describes the analysis of quantitative characteristics of frequent sets and association rules in the posts of Twitter microblogs related to different event discussions. For the analysis, we used a theory of frequent sets,…

Social and Information Networks · Computer Science 2013-10-15 Bohdan Pavlyshenko

Large language models exhibit societal biases associated with demographic information, including race, gender, and others. Endowing such language models with personalities based on demographic data can enable generating opinions that align…

Artificial Intelligence · Computer Science 2024-02-29 Seungjong Sun , Eungu Lee , Dongyan Nan , Xiangying Zhao , Wonbyung Lee , Bernard J. Jansen , Jang Hyun Kim

The Stereotype Content model (SCM) states that we tend to perceive minority groups as cold, incompetent or both. In this paper we adapt existing work to demonstrate that the Stereotype Content model holds for contextualised word embeddings,…

Computation and Language · Computer Science 2022-10-27 Eddie L. Ungless , Amy Rafferty , Hrichika Nag , Björn Ross