中文
相关论文

相关论文: On a construction of a partially non-anticipative …

200 篇论文

Decision-making problems of sequential nature, where decisions made in the past may have an impact on the future, are used to model many practically important applications. In some real-world applications, feedback about a decision is…

机器学习 · 计算机科学 2023-03-02 Ronald C. van den Broek , Rik Litjens , Tobias Sagis , Luc Siecker , Nina Verbeeke , Pratik Gajane

Engineering optimization is typically multiobjective and multidisciplinary with complex constraints, and the solution of such complex problems requires efficient optimization algorithms. Recently, Xin-She Yang proposed a bat-inspired…

最优化与控制 · 数学 2012-03-30 Xin-She Yang

An algorithm is proposed to solve robust control problems constrained by partial differential equations with uncertain coefficients, based on the so-called MG/OPT framework. The levels in this MG/OPT hierarchy correspond to discretization…

数值分析 · 数学 2021-07-21 Andreas Van Barel , Stefan Vandewalle

We consider the repeated prisoner's dilemma (PD). We assume that players make their choices knowing only average payoffs from the previous stages. A player's strategy is a function from the convex hull $\mathfrak{S}$ of the set of payoffs…

最优化与控制 · 数学 2018-05-16 Sławomir Plaskacz , Joanna Zwierzchowska

Blame attribution is one of the key aspects of accountable decision making, as it provides means to quantify the responsibility of an agent for a decision making outcome. In this paper, we study blame attribution in the context of…

人工智能 · 计算机科学 2022-01-26 Stelios Triantafyllou , Adish Singla , Goran Radanovic

A temporally abstract action, or an option, is specified by a policy and a termination condition: the policy guides option behavior, and the termination condition roughly determines its length. Generally, learning with longer options (like…

人工智能 · 计算机科学 2017-12-05 Anna Harutyunyan , Peter Vrancx , Pierre-Luc Bacon , Doina Precup , Ann Nowe

In this paper, we consider a multistage expected utility maximization problem where the decision maker's utility function at each stage depends on historical data and the information on the true utility function is incomplete. To mitigate…

最优化与控制 · 数学 2023-02-22 Jia Liu , Zhiping Chen , Huifu Xu

We consider an approach to constructing a non-anticipating selection of a multivalued mapping; such a problem arises in control theory under conditions of uncertainty. The approach is called "unlocking of predicate" and consists in the…

最优化与控制 · 数学 2017-07-25 D. A. Serkov

In dynamic programming and reinforcement learning, the policy for the sequential decision making of an agent in a stochastic environment is usually determined by expressing the goal as a scalar reward function and seeking a policy that…

人工智能 · 计算机科学 2025-02-26 Simon Dima , Simon Fischer , Jobst Heitzig , Joss Oliver

The assignment of tasks to multiple resources becomes an interesting game theoretic problem, when both the task owner and the resources are strategic. In the classical, nonstrategic setting, where the states of the tasks and resources are…

计算机科学与博弈论 · 计算机科学 2012-02-20 Swaprava Nath , Onno Zoeter , Yadati Narahari , Christopher R. Dance

I formulate and characterize the following two-stage choice behavior. The decision maker is endowed with two preferences. She shortlists all maximal alternatives according to the first preference. If the first preference is decisive, in the…

理论经济学 · 经济学 2020-10-29 Mario Vazquez Corte

Planning and reinforcement learning are two key approaches to sequential decision making. Multi-step approximate real-time dynamic programming, a recently successful algorithm class of which AlphaZero [Silver et al., 2018] is an example,…

人工智能 · 计算机科学 2020-05-18 Thomas M. Moerland , Anna Deichler , Simone Baldi , Joost Broekens , Catholijn M. Jonker

We consider a simple model of imprecise comparisons: there exists some $\delta>0$ such that when a subject is given two elements to compare, if the values of those elements (as perceived by the subject) differ by at least $\delta$, then the…

数据结构与算法 · 计算机科学 2015-01-14 Miklos Ajtai , Vitaly Feldman , Avinatan Hassidim , Jelani Nelson

We study observation-based strategies for partially-observable Markov decision processes (POMDPs) with omega-regular objectives. An observation-based strategy relies on partial information about the history of a play, namely, on the past…

计算机科学中的逻辑 · 计算机科学 2015-05-14 Krishnendu Chatterjee , Laurent Doyen , Thomas A. Henzinger

It is argued that transformation processes (generation rules) showing evidence of a long evolutionary history in universal computing systems can be generalized. The explicit function class $ \Omega $ is defined as follows: "Operators whose…

其他计算机科学 · 计算机科学 2023-04-04 Kazuki Otsuka

A novel model for dynamical traps in intermittent human control is proposed. It describes probabilistic, step-wise transitions between two modes of a subject's behavior - active and passive phases in controlling an object's dynamics - using…

适应与自组织系统 · 物理学 2025-03-04 Vasily Lubashevskiy , Ihor Lubashevsky , Namik Gusein-zade

This paper proposes and studies a general form of dynamic $N$-player non-cooperative games called $\alpha$-potential games, where the change of a player's value function upon her unilateral deviation from her strategy is equal to the change…

最优化与控制 · 数学 2025-04-02 Xin Guo , Xinyu Li , Yufei Zhang

This paper enriches preexisting satisfiability tests for unquantified languages, which in turn augment a fragment of Tarski's elementary algebra with unary real functions possessing a continuous first derivative. Two sorts of individual…

计算机科学中的逻辑 · 计算机科学 2025-07-04 G. Buriola , D. Cantone , G. Cincotti , E. G. Omodeo , G. T. Spartà

Most decision-focused learning work has focused on single stage problems whereas many real-world decision problems are more appropriately modelled using multistage optimisation. In multistage problems contextual information is revealed over…

最优化与控制 · 数学 2025-05-29 Egon Peršak , Miguel F. Anjos

Multistage stochastic optimization problems are oftentimes formulated informally in a pathwise way. These are correct in a discrete setting and suitable when addressing computational challenges, for example. But the pathwise problem…

最优化与控制 · 数学 2021-02-23 Paul Dommel , Alois Pichler
‹ 上一页 1 2 3 10 下一页 ›