中文
相关论文

相关论文: Augmented Utilitarianism for AGI Safety

200 篇论文

Scientific research organizations that are developing and deploying Artificial Intelligence (AI) systems are at the intersection of technological progress and ethical considerations. The push for Responsible AI (RAI) in such institutions…

人工智能 · 计算机科学 2023-12-18 Muneera Bano , Didar Zowghi , Pip Shea , Georgina Ibarra

The operationalization of ethics in the technical practices of artificial intelligence (AI) is facing significant challenges. To address the problem of ineffective implementation of AI ethics, we present our diagnosis, analysis, and…

计算机与社会 · 计算机科学 2025-10-14 Weina Jin , Elise Li Zheng , Ghassan Hamarneh

As AI systems increasingly operate with autonomy and adaptability, the traditional boundaries of moral responsibility in techno-social systems are being challenged. This paper explores the evolving discourse on the delegation of…

计算机与社会 · 计算机科学 2024-11-26 Gordana Dodig-Crnkovic , Gianfranco Basti , Tobias Holstein

While the increased use of AI in the manufacturing sector has been widely noted, there is little understanding on the risks that it may raise in a manufacturing organisation. Although various high level frameworks and definitions have been…

Use of artificial intelligence (AI) in human contexts calls for ethical considerations for the design and development of AI-based systems. However, little knowledge currently exists on how to provide useful and tangible tools that could…

计算机与社会 · 计算机科学 2020-01-23 Ville Vakkuri , Kai-Kristian Kemell , Pekka Abrahamsson

Traditional ethical hacking relies on skilled professionals and time-intensive command management, which limits its scalability and efficiency. To address these challenges, we introduce PenTest++, an AI-augmented system that integrates…

密码学与安全 · 计算机科学 2025-02-14 Haitham S. Al-Sinani , Chris J. Mitchell

Crises in peer review capacity, study replication, and AI-fabricated science have intensified interest in automated tools for assessing scientific research. However, the scientific community has a history of decontextualizing and…

计算机与社会 · 计算机科学 2026-01-16 Carole J. Lee

A series of recent developments points towards auditing as a promising mechanism to bridge the gap between principles and practice in AI ethics. Building on ongoing discussions concerning ethics-based auditing, we offer three contributions.…

计算机与社会 · 计算机科学 2021-05-04 Jakob Mokander , Luciano Floridi

Recent advances in artificial general intelligence (AGI), particularly large language models and creative image generation systems have demonstrated impressive capabilities on diverse tasks spanning the arts and humanities. However, the…

Human feedback is critical for aligning AI systems to human values. As AI capabilities improve and AI is used to tackle more challenging tasks, verifying quality and safety becomes increasingly challenging. This paper explores how we can…

人工智能 · 计算机科学 2025-10-31 Rishub Jain , Sophie Bridgers , Lili Janzer , Rory Greig , Tian Huey Teh , Vladimir Mikulik

As AI systems become increasingly sophisticated, questions about machine consciousness and its ethical implications have moved from fringe speculation to mainstream academic debate. Current ethical frameworks in this domain often implicitly…

计算机与社会 · 计算机科学 2025-12-03 Zhou Ziheng , Haiqiang Dai , Bin Ling , Ying Nian Wu , Demetri Terzopoulos

The field of AI alignment aims to steer AI systems toward human goals, preferences, and ethical principles. Its contributions have been instrumental for improving the output quality, safety, and trustworthiness of today's AI models. This…

人工智能 · 计算机科学 2024-11-26 Robert West , Roland Aydin

Within the current AI ethics discourse, there is a gap in empirical research on understanding how AI practitioners understand ethics and socially organize to operationalize ethical concerns, particularly in the context of AI start-ups. This…

计算机与社会 · 计算机科学 2022-06-22 Mona Sloane , Janina Zakrzewski

Artificial Intelligence (AI) has an increasing impact on all areas of people's livelihoods. A detailed look at existing interdisciplinary and transdisciplinary metrics frameworks could bring new insights and enable practitioners to navigate…

计算机与社会 · 计算机科学 2020-09-16 Marek Havrda , Bogdana Rakova

The complexity of dynamics in AI techniques is already approaching that of complex adaptive systems, thus curtailing the feasibility of formal controllability and reachability analysis in the context of AI safety. It follows that the…

人工智能 · 计算机科学 2018-05-24 Vahid Behzadan , Arslan Munir , Roman V. Yampolskiy

As the deployment of artificial intelligence (AI) is changing many fields and industries, there are concerns about AI systems making decisions and recommendations without adequately considering various ethical aspects, such as…

计算机与社会 · 计算机科学 2023-10-02 Conrad Sanderson , Qinghua Lu , David Douglas , Xiwei Xu , Liming Zhu , Jon Whittle

The AI landscape demands a broad set of legal, ethical, and societal considerations to be accounted for in order to develop ethical AI (eAI) solutions which sustain human values and rights. Currently, a variety of guidelines and a handful…

计算机与社会 · 计算机科学 2021-12-03 Anna Felländer , Jonathan Rebane , Stefan Larsson , Mattias Wiggberg , Fredrik Heintz

The more AI agents are deployed in scenarios with possibly unexpected situations, the more they need to be flexible, adaptive, and creative in achieving the goal we have given them. Thus, a certain level of freedom to choose the best path…

人工智能 · 计算机科学 2018-12-11 Francesca Rossi , Nicholas Mattei

This study focuses on the ethical considerations that a consumer perceives with augmented reality (AR) in the context of smartphone applications. Through a systematic review, this research can provide an understanding and ability for…

人机交互 · 计算机科学 2023-06-14 Nicola J Wood

The emergence of agentic AI marks a new phase in the digital transformation of healthcare. Distinct from conventional generative AI, agentic AI systems are capable of autonomous, goal-directed actions and complex task coordination. They…

计算机与社会 · 计算机科学 2026-02-19 Robert Ranisch , Sabine Salloch