English
Related papers

Related papers: Probing Ethical Framework Representations in Large…

200 papers

Do large language models reason morally, or do they merely sound like they do? We investigate whether LLM responses to moral dilemmas exhibit genuine developmental progression through Kohlberg's stages of moral development, or whether…

Artificial Intelligence · Computer Science 2026-03-24 Aryan Kasat , Smriti Singh , Aman Chadha , Vinija Jain

As large language models (LLMs) increasingly integrate into our daily lives, it becomes crucial to understand their implicit biases and moral tendencies. To address this, we introduce a Moral Foundations LLM dataset (MFD-LLM) grounded in…

Computers and Society · Computer Science 2025-04-10 Monika Jotautaite , Mary Phuong , Chatrik Singh Mangat , Maria Angelica Martinez

Ensuring that Large Language Models (LLMs) return just responses which adhere to societal values is crucial for their broader application. Prior research has shown that LLMs often fail to perform satisfactorily on tasks requiring moral…

Computation and Language · Computer Science 2025-10-09 Guangliang Liu , Zimo Qi , Xitong Zhang , Lei Jiang , Kristen Marie Johnson

With the introduction of ChatGPT, Large Language Models (LLMs) have received enormous attention in healthcare. Despite their potential benefits, researchers have underscored various ethical implications. While individual instances have…

Computers and Society · Computer Science 2024-07-09 Joschka Haltaufderheide , Robert Ranisch

Existing behavioral alignment techniques for Large Language Models (LLMs) often neglect the discrepancy between surface compliance and internal unaligned representations, leaving LLMs vulnerable to long-tail risks. More crucially, we posit…

Computation and Language · Computer Science 2026-03-17 Lingyu Li , Yan Teng , Yingchun Wang

This paper explores the integration of human-like emotions and ethical considerations into Large Language Models (LLMs). We first model eight fundamental human emotions, presented as opposing pairs, and employ collaborative LLMs to…

Computation and Language · Computer Science 2024-06-26 Edward Y. Chang

This work provides an explanatory view of how LLMs can apply moral reasoning to both criticize and defend sexist language. We assessed eight large language models, all of which demonstrated the capability to provide explanations grounded in…

Computation and Language · Computer Science 2024-10-02 Rongchen Guo , Isar Nejadgholi , Hillary Dawkins , Kathleen C. Fraser , Svetlana Kiritchenko

Work on morality in large language models (LLMs) has progressed via constitutional AI, reinforcement learning from human feedback (RLHF) and systematic benchmarking, yet it still lacks tools to connect internal moral representations to…

Human-Computer Interaction · Computer Science 2026-03-25 Gunter Bombaerts

The computer security research community regularly tackles ethical questions. The field of ethics / moral philosophy has for centuries considered what it means to be "morally good" or at least "morally allowed / acceptable". Among…

Cryptography and Security · Computer Science 2023-08-08 Tadayoshi Kohno , Yasemin Acar , Wulf Loh

Large Language Models (LLMs) have shown powerful performance and development prospects and are widely deployed in the real world. However, LLMs can capture social biases from unprocessed training data and propagate the biases to downstream…

Computation and Language · Computer Science 2024-02-22 Yingji Li , Mengnan Du , Rui Song , Xin Wang , Ying Wang

Large language models (LLMs) have rapidly become indispensable tools for acquiring information and supporting human decision-making. However, ensuring that these models uphold fairness across varied contexts is critical to their safe and…

Computers and Society · Computer Science 2026-03-05 Xulang Zhang , Rui Mao , Erik Cambria

With the rise of individual and collaborative networks of autonomous agents, AI is deployed in more key reasoning and decision-making roles. For this reason, ethics-based audits play a pivotal role in the rapidly growing fields of AI safety…

Computers and Society · Computer Science 2024-02-06 Jon Chun , Katherine Elkins

As large language models (LLMs) become more deeply integrated into various sectors, understanding how they make moral judgments has become crucial, particularly in the realm of autonomous driving. This study utilized the Moral Machine…

Computation and Language · Computer Science 2024-02-08 Kazuhiro Takemoto

The integration of Artificial Intelligence (AI) into construction project management (CPM) is accelerating, with Large Language Models (LLMs) emerging as accessible decision-support tools. This study aims to critically evaluate the ethical…

Artificial Intelligence · Computer Science 2025-09-08 Somtochukwu Azie , Yiping Meng

Large Language Models (LLMs) have made unprecedented breakthroughs, yet their increasing integration into everyday life might raise societal risks due to generated unethical content. Despite extensive study on specific issues like bias, the…

Computation and Language · Computer Science 2024-03-05 Shitong Duan , Xiaoyuan Yi , Peng Zhang , Tun Lu , Xing Xie , Ning Gu

Large Language Models (LLMs) have demonstrated impressive capabilities in generating fluent text, as well as tendencies to reproduce undesirable social biases. This study investigates whether LLMs reproduce the moral biases associated with…

Computation and Language · Computer Science 2023-06-21 Gabriel Simmons

Human commonsense understanding of the physical and social world is organized around intuitive theories. These theories support making causal and moral judgments. When something bad happens, we naturally ask: who did what, and why? A rich…

Computation and Language · Computer Science 2023-11-01 Allen Nie , Yuhui Zhang , Atharva Amdekar , Chris Piech , Tatsunori Hashimoto , Tobias Gerstenberg

Large language models (LLMs) increasingly find their way into the most diverse areas of our everyday lives. They indirectly influence people's decisions or opinions through their daily use. Therefore, understanding how and which moral…

Computers and Society · Computer Science 2024-07-23 Karina Vida , Fabian Damken , Anne Lauscher

Language models often misinterpret human intentions due to their handling of ambiguity, a limitation well-recognized in NLP research. While morally clear scenarios are more discernible to LLMs, greater difficulty is encountered in morally…

Computation and Language · Computer Science 2024-10-11 Pranav Senthilkumar , Visshwa Balasubramanian , Prisha Jain , Aneesa Maity , Jonathan Lu , Kevin Zhu

Context: Dark patterns are user interface or other software designs that deceive or manipulate users to do things they would not otherwise do. Even though dark patterns have been under active research for a long time, including particularly…

Software Engineering · Computer Science 2025-03-04 Jukka Ruohonen , Jani Koskinen , Søren Harnow Klausen , Anne Gerdes