English
Related papers

Related papers: Evolving Interpretable Constitutions for Multi-Age…

200 papers

The development of AI agents based on large, open-domain language models (LLMs) has paved the way for the development of general-purpose AI assistants that can support human in tasks such as writing, coding, graphic design, and scientific…

Artificial Intelligence · Computer Science 2025-06-03 Mustafa Mert Çelikok , Saptarashmi Bandyopadhyay , Robert Loftin

Emergent communication in artificial agents has been studied to understand language evolution, as well as to develop artificial systems that learn to communicate with humans. We show that agents performing a cooperative navigation task in…

Machine Learning · Computer Science 2020-07-01 Ivana Kajić , Eser Aygün , Doina Precup

Frontier AI developers now train models against long written behavioral specifications, such as Anthropic's constitution (Anthropic, 2025a) and OpenAI's Model Spec (OpenAI, 2025a), integrated into post-training via methods like character…

Artificial Intelligence · Computer Science 2026-05-26 Arya Jakkli , Senthooran Rajamanoharan , Neel Nanda

Agentic data science (ADS) systems are rapidly improving their capability to autonomously analyze, fit, and interpret data, potentially moving towards a future where agents conduct the vast majority of data-science work. However, current…

Artificial Intelligence · Computer Science 2026-05-06 Chandan Singh , Yan Shuo Tan , Weijia Xu , Zelalem Gero , Weiwei Yang , Michel Galley , Jianfeng Gao

This study introduces "CosmoAgent," an innovative artificial intelligence system that utilizes Large Language Models (LLMs) to simulate complex interactions between human and extraterrestrial civilizations. This paper introduces a…

Computation and Language · Computer Science 2025-06-10 Zhaoqian Xue , Beichen Wang , Suiyuan Zhu , Kai Mei , Hua Tang , Wenyue Hua , Mengnan Du , Yongfeng Zhang

Large language models (LLMs), initially developed for generative AI, are now evolving into agentic AI systems, which make decisions in complex, real-world contexts. Unfortunately, while their generative capabilities are well-documented,…

Artificial Intelligence · Computer Science 2026-04-02 Matthew DosSantos DiSorbo , Harang Ju , Sinan Aral

Large Language Models (LLMs) have increasingly been utilized in social simulations, where they are often guided by carefully crafted instructions to stably exhibit human-like behaviors during simulations. Nevertheless, we doubt the…

Artificial Intelligence · Computer Science 2024-10-29 Zengqing Wu , Run Peng , Shuyuan Zheng , Qianying Liu , Xu Han , Brian Inhyuk Kwon , Makoto Onizuka , Shaojie Tang , Chuan Xiao

Situations of conflict giving rise to social dilemmas are widespread in society and game theory is one major way in which they can be investigated. Starting from the observation that individuals in society interact through networks of…

Physics and Society · Physics 2010-11-24 Enea Pestelacci , Marco Tomassini , Leslie Luthi

The emergence of eusocial species is both very rare in evolutionary history and results in remarkably successful species. By inverting an agent based model, agent rules are discovered that display behaviors characteristic of eusocial…

Physics and Society · Physics 2022-09-09 John C. Stevenson

Contemporary artificial intelligence research has been organized around two dominant ambitions: productivity, which treats AI systems as tools for accelerating work and economic output, and alignment, which focuses on ensuring that…

Artificial Intelligence · Computer Science 2026-03-10 W. Russell Neuman , Chad Coleman

Recent advances in Large Language Models (LLMs) have enabled multi-agent systems that simulate real-world interactions with near-human reasoning. While previous studies have extensively examined biases related to protected attributes such…

Artificial Intelligence · Computer Science 2025-06-03 Min Choi , Keonwoo Kim , Sungwon Chae , Sangyeob Baek

As AI systems move from generating text to accomplishing goals through sustained interaction, the ability to model environment dynamics becomes a central bottleneck. Agents that manipulate objects, navigate software, coordinate with others,…

Modern Reinforcement Learning (RL) algorithms are able to outperform humans in a wide variety of tasks. Multi-agent reinforcement learning (MARL) settings present additional challenges, and successful cooperation in mixed-motive groups of…

Multiagent Systems · Computer Science 2024-06-25 Ram Rachum , Yonatan Nakar , Bill Tomlinson , Nitay Alon , Reuth Mirsky

Recent reports of large language models (LLMs) exhibiting behaviors such as deception, threats, or blackmail are often interpreted as evidence of alignment failure or emergent malign agency. We argue that this interpretation rests on a…

Artificial Intelligence · Computer Science 2026-01-14 Didier Sornette , Sandro Claudio Lera , Ke Wu

Traditional rule-based decision-making methods with interpretable advantage, such as finite state machine, suffer from the jitter or deadlock(JoD) problems in extremely dynamic scenarios. To realize agent swarm confrontation, decision…

Artificial Intelligence · Computer Science 2025-03-12 Zhaoqi Dong , Zhinan Wang , Quanqi Zheng , Bin Xu , Lei Chen , Jinhu Lv

Despite substantial progress of large language models (LLMs) for automatic poetry generation, the generated poetry lacks diversity while the training process differs greatly from human learning. Under the rationale that the learning process…

Computation and Language · Computer Science 2024-09-09 Ran Zhang , Steffen Eger

The universe involves many independent co-learning agents as an ever-evolving part of our observed environment. Yet, in practice, Multi-Agent Reinforcement Learning (MARL) applications are typically constrained to small, homogeneous…

Machine Learning · Computer Science 2025-04-29 Yann Bouteiller , Karthik Soma , Giovanni Beltrame

Law codes and regulations help organise societies for centuries, and as AI systems gain more autonomy, we question how human-agent systems can operate as peers under the same norms, especially when resources are contended. We posit that…

Multiagent Systems · Computer Science 2022-02-24 Alex Raymond , Hatice Gunes , Amanda Prorok

We introduce a black-box interpretability framework that learns a verifiable constitution: a natural language summary of how changes to a prompt affect a model's specific behavior, such as its alignment, correctness, or adherence to…

Machine Learning · Computer Science 2026-02-03 Neha Kalibhat , Zi Wang , Prasoon Bajpai , Drew Proud , Wenjun Zeng , Been Kim , Mani Malek

How cooperation evolves and particularly maintains at a large scale remains an open problem for improving humanity across domains ranging from climate change to pandemic response. To shed light on how behavioral norms can resolve the social…

Physics and Society · Physics 2024-01-25 Brian Mintz , Feng Fu
‹ Prev 1 8 9 10 Next ›