中文
相关论文

相关论文: Formalising the intentional stance 2: a coinductiv…

200 篇论文

Programming languages assume programs directly execute effects. When autonomous systems generate behavior dynamically, this assumption becomes problematic: there is no structural mediation point between deciding to act and acting. We define…

编程语言 · 计算机科学 2026-05-26 Alan L. McCann

A temporally abstract action, or an option, is specified by a policy and a termination condition: the policy guides option behavior, and the termination condition roughly determines its length. Generally, learning with longer options (like…

人工智能 · 计算机科学 2017-12-05 Anna Harutyunyan , Peter Vrancx , Pierre-Luc Bacon , Doina Precup , Ann Nowe

In this work, we present a logical formalism for reasoning about quantum systems in finite dimension. Contrary to the usual approach in quantum logic, our formalism is based classical first-order logic, which allows us to use the tools of…

量子物理 · 物理学 2026-02-19 Olivier Brunet

A treatment policy defines when and what treatments are applied to affect some outcome of interest. Data-driven decision-making requires the ability to predict what happens if a policy is changed. Existing methods that predict how the…

机器学习 · 计算机科学 2023-06-21 Çağlar Hızlı , ST John , Anne Juuti , Tuure Saarinen , Kirsi Pietiläinen , Pekka Marttinen

Two traditional paradigms are often used to describe the behavior of agents in multi-agent complex systems. In the first one, agents are considered to be fully rational and systems are seen as multi-player games. In the second one, agents…

计算机科学与博弈论 · 计算机科学 2016-03-17 Mickael Randour

Decision theorists propose a normative theory of rational choice. Traditionally, they assume that they should provide some constant and invariant principles as criteria for rational decisions, and indirectly, for agents. They seek a…

综合经济学 · 经济学 2021-01-25 Saleh Afroogh

Autonomous vehicles need to be designed to abide by the same rules that humans follow. This is challenging, because traffic rules are fuzzy and not well defined, making them incomprehensible to machines. Satisfaction cannot be incorporated…

机器人学 · 计算机科学 2021-02-08 Klemens Esterle , Luis Gressenbuch , Alois Knoll

Computer simulation provides an automatic and safe way for training robotic control policies to achieve complex tasks such as locomotion. However, a policy trained in simulation usually does not transfer directly to the real hardware due to…

机器学习 · 计算机科学 2018-12-05 Wenhao Yu , C. Karen Liu , Greg Turk

We consider a stochastic control problem where the set of controls is not necessarily convex and the system is governed by a nonlinear backward stochastic differential equation. We establish necessary as well as sufficient conditions of…

概率论 · 数学 2008-12-20 Seid Bahlali

This paper motivates the study of decision theory as necessary for aligning smarter-than-human artificial systems with human interests. We discuss the shortcomings of two standard formulations of decision theory, and demonstrate that they…

人工智能 · 计算机科学 2015-07-09 Nate Soares , Benja Fallenstein

We study a continuous-time stochastic Stackelberg game in which a leader seeks to accomplish a primary objective while inferring a hidden parameter of a rational follower. The follower solves an entropy-regularized tracking problem and…

最优化与控制 · 数学 2025-10-08 Ruimeng Hu , Daniel Ralston , Xu Yang , Haosheng Zhou

This paper introduces a novel causal framework for multi-stage decision-making in natural language action spaces where outcomes are only observed after a sequence of actions. While recent approaches like Proximal Policy Optimization (PPO)…

计算与语言 · 计算机科学 2025-02-26 Bohan Zhang , Yixin Wang , Paramveer S. Dhillon

This article presents a formal model demonstrating that genuine autonomy, the ability of a system to self-regulate and pursue objectives, fundamentally implies computational unpredictability from an external perspective. we establish…

人工智能 · 计算机科学 2025-09-17 Poria Azadi

Probabilistic control design is founded on the principle that a rational agent attempts to match modelled with an arbitrary desired closed-loop system trajectory density. The framework was originally proposed as a tractable alternative to…

机器学习 · 计算机科学 2023-11-16 Tom Lefebvre

We consider a stochastic system whose uncontrolled state dynamics are modelled by a general one-dimensional It\^{o} diffusion. The control effort that can be applied to this system takes the form that is associated with the so-called…

概率论 · 数学 2007-11-15 Andrew J. F. Jack , Timothy C. Johnson , Mihail Zervos

An abstract architecture for idealized multi-agent systems whose behaviour is regulated by normative systems is developed and discussed. Agent choices are determined partially by the preference ordering of possible states and partially by…

计算机科学中的逻辑 · 计算机科学 2007-05-23 Jan Odelstad , Magnus Boman

The safe operation of an autonomous system is a complex endeavor, one pivotal element being its decision-making. Decision-making logic can formally be analyzed using model checking or other formal verification approaches. Yet, the…

多智能体系统 · 计算机科学 2023-10-05 Jan Vermaelen , Tom Holvoet

Temporal point processes have been widely applied to model event sequence data generated by online users. In this paper, we consider the problem of how to design the optimal control policy for point processes, such that the stochastic…

机器学习 · 计算机科学 2017-11-13 Yichen Wang , Grady Williams , Evangelos Theodorou , Le Song

We study the problem of deriving policies, or rules, that when enacted on a complex system, cause a desired outcome. Absent the ability to perform controlled experiments, such rules have to be inferred from past observations of the system's…

机器学习 · 计算机科学 2020-09-09 Kailash Budhathoki , Mario Boley , Jilles Vreeken

Reinforcement learning systems will to a greater and greater extent make decisions that significantly impact the well-being of humans, and it is therefore essential that these systems make decisions that conform to our expectations of…

机器学习 · 计算机科学 2022-05-18 Tue Herlau