English
Related papers

Related papers: The Constitutional Filter: Bayesian Estimation of …

200 papers

Ensuring reliable and rule-compliant behavior of autonomous agents in uncertain environments remains a fundamental challenge in modern robotics. Our work shows how neuro-symbolic systems, which integrate probabilistic, symbolic white-box…

Constitutional AI (CAI) guides LLM behavior using constitutions, but identifying which principles are most effective for model alignment remains an open challenge. We introduce the C3AI framework (\textit{Crafting Constitutions for CAI…

Artificial Intelligence · Computer Science 2025-02-25 Yara Kyrychenko , Ke Zhou , Edyta Bogucka , Daniele Quercia

Constitutional AI is a method to oversee and control LLMs based on a set of rules written in natural language. These rules are typically written by human experts, but could in principle be learned automatically given sufficient training…

Artificial Intelligence · Computer Science 2026-03-18 Rushil Thareja , Gautam Gupta , Francesco Pinto , Nils Lukas

Deep stochastic state-space models enable Bayesian filtering in nonlinear, partially observed systems but typically assume a fixed latent structure. When this assumption is violated, parameter adaptation alone may result in persistent…

Systems and Control · Electrical Eng. & Systems 2026-04-10 Thanana Nuchkrua , Sudchai Boonto , Xiaoqi Liu

Traditional methods for aligning Large Language Models (LLMs), such as Reinforcement Learning from Human Feedback (RLHF) and Direct Preference Optimization (DPO), rely on implicit principles, limiting interpretability. Constitutional AI…

Machine Learning · Computer Science 2025-04-01 Carl-Leander Henneking , Claas Beger

There is growing consensus that language model (LM) developers should not be the sole deciders of LM behavior, creating a need for methods that enable the broader public to collectively shape the behavior of LM systems that affect them. To…

Artificial Intelligence · Computer Science 2024-06-13 Saffron Huang , Divya Siddarth , Liane Lovitt , Thomas I. Liao , Esin Durmus , Alex Tamkin , Deep Ganguli

We are increasingly subjected to the power of AI authorities. As AI decisions become inescapable, entering domains such as healthcare, education, and law, we must confront a vital question: how can we ensure AI systems have the legitimacy…

Computers and Society · Computer Science 2025-05-15 Gilad Abiri

Large language models increasingly function as artificial reasoners: they evaluate arguments, assign credibility, and express confidence. Yet their belief-forming behavior is governed by implicit, uninspected epistemic policies. This paper…

Artificial Intelligence · Computer Science 2026-04-23 Michele Loi

Agentic AI systems, possessing capabilities for autonomous planning and action, show great potential across diverse domains. However, their practical deployment is hindered by challenges in aligning their behavior with varied human values,…

Artificial Intelligence · Computer Science 2025-08-12 Nell Watson , Ahmed Amer , Evan Harris , Preeti Ravindra , Shujun Zhang

The Quantum Fisher information (QFI) quantifies the ultimate precision of estimating a parameter from a quantum state, and can be regarded as a reliability measure of a quantum system as a quantum sensor. However, estimation of the QFI for…

Quantum Physics · Physics 2022-02-03 Jacob L. Beckey , M. Cerezo , Akira Sone , Patrick J. Coles

As language models continue to grow larger, the cost of acquiring high-quality training data has increased significantly. Collecting human feedback is both expensive and time-consuming, and manual labels can be noisy, leading to an…

Artificial Intelligence · Computer Science 2025-04-08 Xue Zhang

Reasoning on large and complex real-world models is a computationally difficult task, yet one that is required for effective use of many AI applications. A plethora of inference algorithms have been developed that work well on specific…

Artificial Intelligence · Computer Science 2016-06-13 Avi Pfeffer , Brian Ruttenberg , William Kretschmer

The robust estimation of dynamically changing features, such as the position of prey, is one of the hallmarks of perception. On an abstract, algorithmic level, nonlinear Bayesian filtering, i.e. the estimation of temporally changing signals…

Neurons and Cognition · Quantitative Biology 2022-01-05 Anna Kutschireiter , Simone Carlo Surace , Henning Sprekeler , Jean-Pascal Pfister

This paper presents a computational account of how legal norms can influence the behavior of artificial intelligence (AI) agents, grounded in the active inference framework (AIF) that is informed by principles of economic legal analysis…

Computers and Society · Computer Science 2025-11-25 Axel Constant , Mahault Albarracin , Karl J. Friston

As computational power has continued to increase, and sensors have become more accurate, the corresponding advent of systems that are at once cognitive and immersive has arrived. These \textit{cognitive and immersive systems} (CAISs) fall…

Artificial Intelligence · Computer Science 2024-06-10 Matthew Peveler , Naveen Sundar Govindarajulu , Selmer Bringsjord , Atriya Sen , Biplav Srivastava , Kartik Talamadupula , Hui Su

Cooperation is often implicitly assumed when learning from other agents. Cooperation implies that the agent selecting the data, and the agent learning from the data, have the same goal, that the learner infer the intended hypothesis. Recent…

Machine Learning · Computer Science 2020-07-02 Junqi Wang , Pei Wang , Patrick Shafto

We are currently unable to specify human goals and societal values in a way that reliably directs AI behavior. Law-making and legal interpretation form a computational engine that converts opaque human values into legible directives. "Law…

Computers and Society · Computer Science 2023-05-17 John J. Nay

Scientific Artificial Intelligence (AI) applications require models that deliver trustworthy uncertainty estimates while respecting domain constraints. Existing uncertainty quantification methods lack mechanisms to incorporate symbolic…

Machine Learning · Computer Science 2026-01-21 Shahnawaz Alam , Mohammed Mudassir Uddin , Mohammed Kaif Pasha

With the rapid development of large language models (LLMs), aligning LLMs with human values and societal norms to ensure their reliability and safety has become crucial. Reinforcement learning with human feedback (RLHF) and Constitutional…

Computation and Language · Computer Science 2024-03-28 Xiusi Chen , Hongzhi Wen , Sreyashi Nag , Chen Luo , Qingyu Yin , Ruirui Li , Zheng Li , Wei Wang
‹ Prev 1 2 3 10 Next ›