English
Related papers

Related papers: Specific versus General Principles for Constitutio…

200 papers

There is substantial concern about the ability of advanced artificial intelligence to influence people's behaviour. A rapidly growing body of research has found that AI can produce large persuasive effects on people's attitudes, but whether…

Computers and Society · Computer Science 2026-04-13 Kobi Hackenburg , Luke Hewitt , Caroline Wagner , Ben M. Tappin , Christopher Summerfield

Conversational AI is rapidly becoming a primary interface for information seeking and decision making, yet most systems still assume idealized users. In practice, human reasoning is bounded by limited attention, uneven knowledge, and…

Emerging Technologies · Computer Science 2026-01-21 Jiqun Liu

As humans increasingly rely on multiround conversational AI for high stakes decisions, principled frameworks are needed to ensure such interactions reliably improve decision quality. We adopt a human centric view governed by two principles:…

Machine Learning · Computer Science 2026-02-25 Sima Noorani , Shayan Kiyani , Hamed Hassani , George Pappas

In many contexts, lying -- the use of verbal falsehoods to deceive -- is harmful. While lying has traditionally been a human affair, AI systems that make sophisticated verbal statements are becoming increasingly prevalent. This raises the…

Computers and Society · Computer Science 2021-10-14 Owain Evans , Owen Cotton-Barratt , Lukas Finnveden , Adam Bales , Avital Balwit , Peter Wills , Luca Righetti , William Saunders

Existing methods for controlling language models, such as RLHF and Constitutional AI, involve determining which LLM behaviors are desirable and training them into a language model. However, in many cases, it is desirable for LLMs to be…

Computation and Language · Computer Science 2024-02-14 Louis Castricato , Nathan Lile , Suraj Anand , Hailey Schoelkopf , Siddharth Verma , Stella Biderman

The impact of Artificial Intelligence does not depend only on fundamental research and technological developments, but for a large part on how these systems are introduced into society and used in everyday situations. AI is changing the way…

Computers and Society · Computer Science 2022-05-24 Virginia Dignum

This paper challenges the assumption that courts should grant First Amendment protections to outputs from large generative AI models, such as GPT-4 and Gemini. We argue that because these models lack intentionality, their outputs do not…

Computers and Society · Computer Science 2025-06-06 David Atkinson , Jena D. Hwang , Jacob Morrison

Reinforcement Learning from AI Feedback (RLAIF) enables language models to improve by training on their own preference judgments, yet no theoretical account explains why this self-improvement seemingly works for value learning. We propose…

Machine Learning · Computer Science 2026-03-04 Robin Young

The article further develops and formalizes a theory of friendly dialogue in an AI System of Dr. Watson type, as proposed in our previous publication[4],[19]. The main principle of this type of AI is to guide the user toward a solution in a…

Artificial Intelligence · Computer Science 2024-07-31 Saveli Goldberg , Vladimir Sluchak

AI is powerful, but it can make choices that result in objective errors, contextually inappropriate outputs, and disliked options. We need AI-resilient interfaces that help people be resilient to the AI choices that are not right, or not…

Human-Computer Interaction · Computer Science 2024-05-15 Elena L. Glassman , Ziwei Gu , Jonathan K. Kummerfeld

Traditionally, the way one evaluates the performance of an Artificial Intelligence (AI) system is via a comparison to human performance in specific tasks, treating humans as a reference for high-level cognition. However, these comparisons…

Artificial Intelligence · Computer Science 2019-11-25 Camilo M. Signorelli , Xerxes D. Arsiwalla

One of today's most significant societal challenges is building AI systems whose behaviour, or the behaviour it enables within communities of interacting agents (human and artificial), aligns with human values. To address this challenge, we…

Artificial Intelligence · Computer Science 2026-02-09 Nardine Osman , Mark d'Inverno

As AI systems increasingly permeate everyday life, designers and developers face mounting pressure to balance innovation with ethical design choices. To date, the operationalisation of AI ethics has predominantly depended on frameworks that…

Human-Computer Interaction · Computer Science 2025-09-18 Benjamin J. Carroll , Jianlong Zhou , Paul F. Burke , Sabine Ammon

Moral judgements form the foundation of human social behavior and societal systems. While Artificial Intelligence chatbots increasingly serve as personal advisors, their influence on moral judgments remains largely unexplored. Here, we…

Artificial Intelligence · Computer Science 2026-04-24 Yue Teng , Qianer Zhong , Kim Mai Tich Nguyen Thordsen , Christian Montag , Benjamin Becker

Artificial Intelligence (AI) has been used extensively in automatic decision making in a broad variety of scenarios, ranging from credit ratings for loans to recommendations of movies. Traditional design guidelines for AI models focus…

Artificial Intelligence · Computer Science 2018-09-27 Marisa Vasconcelos , Carlos Cardonha , Bernardo Gonçalves

Although artificial intelligence (AI) is solving real-world challenges and transforming industries, there are serious concerns about its ability to behave and make decisions in a responsible way. Many AI ethics principles and guidelines for…

Artificial Intelligence · Computer Science 2022-07-22 Qinghua Lu , Liming Zhu , Xiwei Xu , Jon Whittle , David Douglas , Conrad Sanderson

Multi-agent AI systems need behavioral constitutions, but it is unresolved whether such rules should emerge internally through agent self-governance or be discovered externally through optimization. We present the first controlled…

Multiagent Systems · Computer Science 2026-05-12 Hershraj Niranjani , Ujwal Kumar , Phan Xuan Tan

Transformative AI systems may pose unprecedented catastrophic risks, but the U.S. Constitution places significant constraints on the government's ability to govern this technology. This paper examines how the First Amendment, administrative…

Computers and Society · Computer Science 2025-12-19 Alex Mark , Aaron Scher

AI ethics is an emerging field with multiple, competing narratives about how to best solve the problem of building human values into machines. Two major approaches are focused on bias and compliance, respectively. But neither of these ideas…

Artificial Intelligence · Computer Science 2023-02-24 Thomas Krendl Gilbert , Megan Welle Brozek , Andrew Brozek

As Generative AI systems increasingly engage in long-term, personal, and relational interactions, human-AI engagements are becoming significantly complex, making them more challenging to understand and govern. These Interactive AI systems…

Computers and Society · Computer Science 2025-08-26 Yulu Pi , Cagatay Turkay , Daniel Bogiatzis-Gibbons