English
Related papers

Related papers: NeuroAI for AI Safety

200 papers

Artificial Intelligence (AI) is revolutionizing assistive technologies. It offers innovative solutions to enhance the quality of life for individuals with visual impairments. This review examines the development, applications, and impact of…

Human-Computer Interaction · Computer Science 2025-03-21 Prudhvi Naayini , Praveen Kumar Myakala , Chiranjeevi Bura , Anil Kumar Jonnalagadda , Srikanth Kamatala

Self-improvement is a goal currently exciting the field of AI, but is fraught with danger, and may take time to fully achieve. We advocate that a more achievable and better goal for humanity is to maximize co-improvement: collaboration…

Artificial Intelligence · Computer Science 2025-12-16 Jason Weston , Jakob Foerster

Current AI governance frameworks, including regulatory benchmarks for accuracy, latency, and energy efficiency, are built for static, centrally trained artificial neural networks on von Neumann hardware. NeuroAI systems, embodied in…

Emerging Technologies · Computer Science 2026-02-06 Afifah Kashif , Abdul Muhsin Hameed , Asim Iqbal

This paper critically examines the evolving ethical and regulatory challenges posed by the integration of artificial intelligence (AI) in cybersecurity. We trace the historical development of AI regulation, highlighting major milestones…

Cryptography and Security · Computer Science 2025-01-22 Vikram Kulothungan

We present an overview of the literature on trust in AI and AI trustworthiness and argue for the need to distinguish these concepts more clearly and to gather more empirically evidence on what contributes to people s trusting behaviours. We…

Artificial Intelligence · Computer Science 2023-09-20 Andreas Duenser , David M. Douglas

This paper presents a conceptual and operational framework for developing and operating safe and trustworthy AI agents based on a Three-Pillar Model grounded in transparency, accountability, and trustworthiness. Building on prior work in…

Computers and Society · Computer Science 2026-01-13 Edward C. Cheng , Jeshua Cheng , Alice Siu

Cloud security concerns have been greatly realized in recent years due to the increase of complicated threats in the computing world. Many traditional solutions do not work well in real-time to detect or prevent more complex threats.…

Cryptography and Security · Computer Science 2025-05-08 Shamnad Mohamed Shaffi , Sunish Vengathattil , Jezeena Nikarthil Sidhick , Resmi Vijayan

Artificial Intelligence (AI) is poised to transform healthcare delivery through revolutionary advances in clinical decision support and diagnostic capabilities. While human expertise remains foundational to medical practice, AI-powered…

Every AI system is deployed by a human organization. In high risk applications, the combined human plus AI system must function as a high-reliability organization in order to avoid catastrophic errors. This short note reviews the properties…

Artificial Intelligence · Computer Science 2018-11-28 Thomas G. Dietterich

The increasing use of AI technologies has led to increasing AI incidents, posing risks and causing harm to individuals, organizations, and society. This study recognizes and addresses the lack of standardized protocols for reliably and…

Computers and Society · Computer Science 2025-01-28 Avinash Agarwal , Manisha J Nene

AI is displacing tasks, mediating high-stakes decisions, and flooding communication with synthetic content, unsettling work, identity, and social trust. We argue that the decisive human countermeasure is resilience. We define resilience…

Computers and Society · Computer Science 2025-10-30 Shaoshan Liu , Anina Schwarzenbach , Yiyu Shi

Memory serves as the pivotal nexus bridging past and future, providing both humans and AI systems with invaluable concepts and experience to navigate complex tasks. Recent research on autonomous agents has increasingly focused on designing…

Computation and Language · Computer Science 2025-12-30 Jiafeng Liang , Hao Li , Chang Li , Jiaqi Zhou , Shixin Jiang , Zekun Wang , Changkai Ji , Zhihao Zhu , Runxuan Liu , Tao Ren , Jinlan Fu , See-Kiong Ng , Xia Liang , Ming Liu , Bing Qin

Rapidly improving AI capabilities and autonomy hold significant promise of transformation, but are also driving vigorous debate on how to ensure that AI is safe, i.e., trustworthy, reliable, and secure. Building a trusted ecosystem is…

The human brain is the substrate for human intelligence. By simulating the human brain, artificial intelligence builds computational models that have learning capabilities and perform intelligent tasks approaching the human level. Deep…

Neurons and Cognition · Quantitative Biology 2024-02-20 Barco Jie You

Safety and responsibility evaluations of advanced AI models are a critical but developing field of research and practice. In the development of Google DeepMind's advanced AI models, we innovated on and applied a broad set of approaches to…

Major AI ethics guidelines and laws, including the EU AI Act, call for effective human oversight, but do not define it as a distinct and developable capacity. This paper introduces human oversight as a well-being capacity, situated within…

Computers and Society · Computer Science 2025-12-17 Yao Xie , Walter Cullen

In the last years, AI systems, in particular neural networks, have seen a tremendous increase in performance, and they are now used in a broad range of applications. Unlike classical symbolic AI systems, neural networks are trained using…

Computer Vision and Pattern Recognition · Computer Science 2021-08-16 Christian Berghoff , Pavol Bielik , Matthias Neu , Petar Tsankov , Arndt von Twickel

Artificial Intelligence has made remarkable advancements in recent years, primarily driven by increasingly large deep learning models. However, achieving true Artificial General Intelligence (AGI) demands fundamentally new architectures…

Artificial Intelligence · Computer Science 2025-04-30 Rajeev Gupta , Suhani Gupta , Ronak Parikh , Divya Gupta , Amir Javaheri , Jairaj Singh Shaktawat

Power is a key concept in AI safety: power-seeking as an instrumental goal, sudden or gradual disempowerment of humans, power balance in human-AI interaction and international AI governance. At the same time, power as the ability to pursue…

Artificial Intelligence · Computer Science 2025-08-06 Jobst Heitzig , Ram Potham

We envision "AI scientists" as systems capable of skeptical learning and reasoning that empower biomedical research through collaborative agents that integrate AI models and biomedical tools with experimental platforms. Rather than taking…