English
Related papers

Related papers: Formalising the intentional stance 2: a coinductiv…

200 papers

In recent papers it has been suggested that human locomotion may be modeled as an inverse optimal control problem. In this paradigm, the trajectories are assumed to be solutions of an optimal control problem that has to be determined. We…

Optimization and Control · Mathematics 2010-07-26 Yacine Chitour , Frédéric Jean , Paolo Mason

Striking a balance between efficiency and transparent motion is a core challenge in human-robot collaboration, as highly expressive movements often incur unnecessary time and energy costs. In collaborative environments, legibility allows a…

Robotics · Computer Science 2026-05-07 Adrien Jacquet Crétides , Mouad Abrini , Hamed Rahimi , Mohamed Chetouani

The aim of the present paper is to provide criteria for a central bank of how to choose among different monetary-policy rules when caring about a number of policy targets such as the output gap and expected inflation. Special attention is…

General Economics · Economics 2020-12-08 Jean-Bernard Chatelain , Kirsten Ralf

Stochastic processes offer a flexible mathematical formalism to model and reason about systems. Most analysis tools, however, start from the premises that models are fully specified, so that any parameters controlling the system's dynamics…

Systems and Control · Computer Science 2017-01-11 Luca Bortolussi , Guido Sanguinetti

In Willems' behavioral systems theory, a dynamical system is identified with the set of all trajectories compatible with its laws of motion. In the linear time-invariant setting this trajectory set is a linear subspace, and its algebraic…

Optimization and Control · Mathematics 2026-05-08 Victor M. Preciado

We investigate the performance of m-th order consensus systems with stochastic external perturbations, where a subset of leader nodes incorporates absolute information into their control laws. The system performance is measured by its…

Optimization and Control · Mathematics 2019-06-13 Erika Mackin , Stacy Patterson

The iterated prisoner's dilemma is a game that produces many counter-intuitive and complex behaviors in a social environment, based on very simple basic rules. It illustrates that cooperation can be a good thing even in a competitive world,…

Computer Science and Game Theory · Computer Science 2020-09-07 Robert Prentner

This paper considers a half-duplex scenario where an interferer behaves according to a parametric model but the values of the model parameters are unknown. We explore the necessary number of sensing steps to gather sufficient knowledge…

Information Theory · Computer Science 2024-10-11 Vincent Corlay , Jean-Christophe Sibel , Nicolas Gresset

In this article, we propose to use the formalism of quantum mechanics to describe and explain the so-called "abnormal" behaviour of agents in certain decision or choice contexts. The basic idea is to postulate that the preferences of these…

Physics and Society · Physics 2024-12-04 Herve Zwirn

Causal reasoning has been an indispensable capability for humans and other intelligent animals to interact with the physical world. In this work, we propose to endow an artificial agent with the capability of causal reasoning for completing…

Machine Learning · Computer Science 2019-10-07 Suraj Nair , Yuke Zhu , Silvio Savarese , Li Fei-Fei

As learned control policies become increasingly common in autonomous systems, there is increasing need to ensure that they are interpretable and can be checked by human stakeholders. Formal specifications have been proposed as ways to…

Human-Computer Interaction · Computer Science 2024-07-04 Isabelle Hurley , Rohan Paleja , Ashley Suh , Jaime D. Peña , Ho Chit Siu

What is the difference between goal-directed and habitual behavior? We propose a novel computational framework of decision making with Bayesian inference, in which everything is integrated as an entire neural network model. The model learns…

Machine Learning · Computer Science 2021-06-23 Dongqi Han , Kenji Doya , Jun Tani

In the theory of dynamic programming, an optimal policy is a policy whose lifetime value dominates that of all other policies from every possible initial condition in the state space. This raises a natural question: when does optimality…

Optimization and Control · Mathematics 2025-05-13 John Stachurski , Jingni Yang , Ziyue Yang

Modern large language models (LLMs) employ diverse logical inference mechanisms for reasoning, making the strategic optimization of these approaches critical for advancing their capabilities. This paper systematically investigate the…

Computation and Language · Computer Science 2025-09-18 Tianshi Zheng , Jiayang Cheng , Chunyang Li , Haochen Shi , Zihao Wang , Jiaxin Bai , Yangqiu Song , Ginny Y. Wong , Simon See

The proliferation of artificial intelligence is increasingly dependent on model understanding. Understanding demands both an interpretation - a human reasoning about a model's behavior - and an explanation - a symbolic representation of the…

Machine Learning · Computer Science 2022-08-30 Charl Maree , Christian W. Omlin

Adaptive systems -- such as a biological organism gaining survival advantage, an autonomous robot executing a functional task, or a motor protein transporting intracellular nutrients -- must model the regularities and stochasticity in their…

Statistical Mechanics · Physics 2021-04-13 A. B. Boyd , J. P. Crutchfield , M. Gu

The idea is advanced that self-organization in complex systems can be treated as decision making (as it is performed by humans) and, vice versa, decision making is nothing but a kind of self-organization in the decision maker nervous…

Adaptation and Self-Organizing Systems · Physics 2014-08-08 V. I. Yukalov , D. Sornette

Some researchers speculate that intelligent reinforcement learning (RL) agents would be incentivized to seek resources and power in pursuit of their objectives. Other researchers point out that RL agents need not have human-like…

Artificial Intelligence · Computer Science 2023-01-31 Alexander Matt Turner , Logan Smith , Rohin Shah , Andrew Critch , Prasad Tadepalli

This paper considers the problem of designing time-dependent, real-time control policies for controllable nonlinear diffusion processes, with the goal of obtaining maximally-informative observations about parameters of interest. More…

Methodology · Statistics 2014-10-16 Giles Hooker , Kevin K. Lin , Bruce Rogers

The primary goal of reinforcement learning is to develop decision-making policies that prioritize optimal performance, frequently without considering safety. In contrast, safe reinforcement learning seeks to reduce or avoid unsafe behavior.…

Machine Learning · Computer Science 2025-06-17 Zahra Shahrooei , Ali Baheri