中文
相关论文

相关论文: Taxonomizing Representational Harms using Speech A…

200 篇论文

Organizations worldwide that rely on data-driven approaches regularly employ forecasting methods to enhance their planning and decision-making processes. While extensive research has examined the harms associated with traditional machine…

其他统计学 · 统计学 2025-03-14 Bahman Rostami-Tabar , Travis Greene , Galit Shmueli , Rob J. Hyndman

Drawing on work spanning economics, public health, education, sociology, and law, I formalize theoretically what makes systemic discrimination "systemic." Injustices do not occur in isolation, but within a complex system of interdependent…

综合经济学 · 经济学 2024-03-19 David B. McMillon

Recent advances in Bayesian probability theory and its application to cognitive science in combination with the development of a new generation of computational tools and methods for probabilistic computation have led to a 'probabilistic…

计算与语言 · 计算机科学 2025-09-29 Christoph Unger , Hendrik Buschmeier

Hate speech causes widespread and deep-seated societal issues. Proper enforcement of hate speech laws is key for protecting groups of people against harmful and discriminatory language. However, determining what constitutes hate speech is a…

计算与语言 · 计算机科学 2023-11-03 Chu Fei Luo , Rohan Bhambhoria , Xiaodan Zhu , Samuel Dahan

This paper introduces a collaborative, human-centred taxonomy of AI, algorithmic and automation harms. We argue that existing taxonomies, while valuable, can be narrow, unclear, typically cater to practitioners and government, and often…

As autonomous systems rapidly become ubiquitous, there is a growing need for a legal and regulatory framework to address when and how such a system harms someone. There have been several attempts within the philosophy literature to define…

人工智能 · 计算机科学 2023-01-20 Sander Beckers , Hana Chockler , Joseph Y. Halpern

To facilitate the measurement of representational harms caused by large language model (LLM)-based systems, the NLP research community has produced and made publicly available numerous measurement instruments, including tools, datasets,…

As we increasingly delegate decision-making to algorithms, whether directly or indirectly, important questions emerge in circumstances where those decisions have direct consequences for individual rights and personal opportunities, as well…

计算机与社会 · 计算机科学 2019-05-01 Teresa Scantamburlo , Andrew Charlesworth , Nello Cristianini

As large language models (LLMs) achieve advanced persuasive capabilities, concerns about their potential risks have grown. The EU AI Act prohibits AI systems that use manipulative or deceptive techniques to undermine informed…

计算机与社会 · 计算机科学 2025-05-20 Haein Kong

Hate speech has grown significantly on social media, causing serious consequences for victims of all demographics. Despite much attention being paid to characterize and detect discriminatory speech, most work has focused on explicit or…

This paper focuses on a referring expression generation (REG) task in which the aim is to pick out an object in a complex visual scene. One common theoretical approach to this problem is to model the task as a two-agent cooperative scheme…

计算与语言 · 计算机科学 2022-05-17 Hieu Le , Taufiq Daryanto , Fabian Zhafransyah , Derry Wijaya , Elizabeth Coppock , Sang Chin

As machine learning applications proliferate, we need an understanding of their potential for harm. However, current fairness metrics are rarely grounded in human psychological experiences of harm. Drawing on the social psychology of…

计算机与社会 · 计算机科学 2025-05-27 Angelina Wang , Xuechunzi Bai , Solon Barocas , Su Lin Blodgett

Counterspeech, i.e., responses to counteract potential harms of hateful speech, has become an increasingly popular solution to address online hate speech without censorship. However, properly countering hateful language requires countering…

计算与语言 · 计算机科学 2023-11-02 Jimin Mun , Emily Allaway , Akhila Yerukola , Laura Vianna , Sarah-Jane Leslie , Maarten Sap

This work proposes a contextualised detection framework for implicitly hateful speech, implemented as a multi-agent system comprising a central Moderator Agent and dynamically constructed Community Agents representing specific demographic…

计算与语言 · 计算机科学 2026-01-28 Ewelina Gajewska , Katarzyna Budzynska , Jarosław A Chudziak

General-purpose safety benchmarks for large language models do not adequately evaluate disability-related harms. We introduce DisaBench: a taxonomy of twelve disability harm categories co-created with people with disabilities and red…

人工智能 · 计算机科学 2026-05-14 Eugenia Kim , Ioana Tanase , Christina Mallon

Considering the large amount of content created online by the minute, slang-aware automatic tools are critically needed to promote social good, and assist policymakers and moderators in restricting the spread of offensive language, abuse,…

计算与语言 · 计算机科学 2023-02-02 Aravinda Kolla , Filip Ilievski , Hông-Ân Sandlin , Alain Mermoud

Social choice theory is the study of preference aggregation across a population, used both in mechanism design for human agents and in the democratic alignment of language models. In this study, we propose the representative social choice…

机器学习 · 计算机科学 2025-11-03 Tianyi Qiu

The growing influence of Artificial Intelligence (AI) systems on decision-making in critical domains has exposed their potential to cause significant harms, often rooted in biases embedded across the AI lifecycle. While existing frameworks…

计算机与社会 · 计算机科学 2025-12-04 Nicoleta Tantalaki , Sophia Vei , Athena Vakali

Gender is widely discussed in the context of language tasks and when examining the stereotypes propagated by language models. However, current discussions primarily treat gender as binary, which can perpetuate harms such as the cyclical…

计算与语言 · 计算机科学 2021-09-14 Sunipa Dev , Masoud Monajatipoor , Anaelia Ovalle , Arjun Subramonian , Jeff M Phillips , Kai-Wei Chang

In this paper, we show how game-theoretic work on conversation combined with a theory of discourse structure provides a framework for studying interpretive bias. Interpretive bias is an essential feature of learning and understanding but…

计算与语言 · 计算机科学 2018-07-02 Nicholas Asher , Soumya Paul