中文
相关论文

相关论文: Formalising the intentional stance 2: a coinductiv…

200 篇论文

In dynamic programming and reinforcement learning, the policy for the sequential decision making of an agent in a stochastic environment is usually determined by expressing the goal as a scalar reward function and seeking a policy that…

人工智能 · 计算机科学 2025-02-26 Simon Dima , Simon Fischer , Jobst Heitzig , Joss Oliver

Traditionally, in Programming-by-example (PBE) the goal is to synthesize a program from a small set of input-output examples. Lately, PBE has gained traction as a few-shot reasoning benchmark, relaxing the requirement to produce a program…

编程语言 · 计算机科学 2026-03-17 Janis Zenkner , Tobias Sesterhenn , Christian Bartelt

Understanding a \textit{reinforcement learning} policy, which guides state-to-action mappings to maximize rewards, necessitates an accompanying explanation for human comprehension. In this paper, we introduce a set of \textit{linear…

人工智能 · 计算机科学 2025-05-01 Mikihisa Yuasa , Huy T. Tran , Ramavarapu S. Sreenivas

The scheme of online optimization as a feedback controller is widely used to steer the states of a physical system to the optimal solution of a predefined optimization problem. Such methods focus on regulating the physical states to the…

系统与控制 · 电气工程与系统科学 2025-08-05 Ming Li , Zhaojian Wang , Feng Liu , Ming Cao , Bo Yang

People are often confronted with problems whose complexity exceeds their cognitive capacities. To deal with this complexity, individuals and managers can break complex problems down into a series of subgoals. Which subgoals are most…

人工智能 · 计算机科学 2023-02-07 Nishad Singhi , Florian Mohnert , Ben Prystawski , Falk Lieder

Autonomous agents operating in public spaces must consider how their behaviors might affect the humans around them, even when not directly interacting with them. To this end, it is often beneficial to be predictable and appear naturalistic.…

多智能体系统 · 计算机科学 2025-05-06 Hamzah I. Khan , David Fridovich-Keil

While reinforcement learning algorithms provide automated acquisition of optimal policies, practical application of such methods requires a number of design decisions, such as manually designing reward functions that not only define the…

机器学习 · 计算机科学 2022-12-29 Tim G. J. Rudner , Vitchyr H. Pong , Rowan McAllister , Yarin Gal , Sergey Levine

This paper introduces a continuous-time constrained nonlinear control scheme which implements a model predictive control strategy as a continuous-time dynamic system. The approach is based on the idea that the solution of the optimal…

系统与控制 · 计算机科学 2017-09-20 Marco M. Nicotra , Dominic Liao-McPherson , Ilya V. Kolmanovsky

Behavior constrained policy optimization has been demonstrated to be a successful paradigm for tackling Offline Reinforcement Learning. By exploiting historical transitions, a policy is trained to maximize a learned value function while…

机器学习 · 计算机科学 2023-07-25 Jiachen Li , Edwin Zhang , Ming Yin , Qinxun Bai , Yu-Xiang Wang , William Yang Wang

In this paper we propose to use elements of the mathematical formalism of Quantum Mechanics to capture the idea that agents' preferences, in addition to being typically uncertain, can also be indeterminate. They are determined (i.e.,…

物理与社会 · 物理学 2007-09-03 Ariane Lambert-Mogiliansky , Shmuel Zamir , Herve Zwirn

There are two fundamentally different approaches to specifying and verifying properties of systems. The logical approach makes use of specifications given as formulae of temporal or modal logics and relies on efficient model checking…

计算机科学中的逻辑 · 计算机科学 2013-06-05 Nikola Beneš , Benoît Delahaye , Uli Fahrenberg , Jan Křetínský , Axel Legay

While many multiagent algorithms are designed for homogeneous systems (i.e. all agents are identical), there are important applications which require an agent to coordinate its actions without knowing a priori how the other agents behave.…

人工智能 · 计算机科学 2019-07-17 Stefano V. Albrecht , Subramanian Ramamoorthy

We present a risk-aware formalism for evaluating system trajectories in the presence of uncertain interactions between the system and its environment. The proposed formalism supports reasoning under uncertainty and systematically handles…

系统与控制 · 电气工程与系统科学 2026-04-28 Tichakorn Wongpiromsarn

Game theory serves as a powerful tool for distributed optimization in multi-agent systems in different applications. In this paper we consider multi-agent systems that can be modeled by means of potential games whose potential function…

最优化与控制 · 数学 2018-04-13 Tatiana Tatarenko

We study the problem of identifying the policy space of a learning agent, having access to a set of demonstrations generated by its optimal policy. We introduce an approach based on statistical testing to identify the set of policy…

机器学习 · 计算机科学 2019-09-10 Alberto Maria Metelli , Guglielmo Manneschi , Marcello Restelli

Extremal principles are fundamental in our interpretation of phenomena in nature. One of the best known examples is the second law of thermodynamics, governing most physical and chemical systems and stating the continuous increase of…

统计力学 · 物理学 2007-05-23 Dirk Helbing , Tamas Vicsek

We consider the synthesis of control policies for probabilistic systems, modeled by Markov decision processes, operating in partially known environments with temporal logic specifications. The environment is modeled by a set of Markov…

计算机科学中的逻辑 · 计算机科学 2012-03-07 Tichakorn Wongpiromsarn , Emilio Frazzoli

A decision maker typically (i) incorporates training data to learn about the relative effectiveness of treatments, and (ii) chooses an implementation mechanism that implies an ``optimal'' predicted outcome distribution according to some…

计量经济学 · 经济学 2025-05-29 Anders Bredahl Kock , David Preinerstorfer

In Reinforcement Learning interpretability generally means to provide insight into the agent's mechanisms such that its decisions are understandable by an expert upon inspection. This definition, with the resulting methods from the…

人工智能 · 计算机科学 2022-03-10 Michele Persiani , Thomas Hellström

A central concept in active inference is that the internal states of a physical system parametrise probability measures over states of the external world. These can be seen as an agent's beliefs, expressed as a Bayesian prior or posterior.…

人工智能 · 计算机科学 2021-12-28 Nathaniel Virgo , Martin Biehl , Simon McGregor