中文
相关论文

相关论文: Strategic control for a Boltzmann like decision-ma…

200 篇论文

We consider a ubiquitous scenario in the Internet economy when individual decision-makers (henceforth, agents) both produce and consume information as they make strategic choices in an uncertain environment. This creates a three-way…

计算机科学与博弈论 · 计算机科学 2021-04-09 Yishay Mansour , Aleksandrs Slivkins , Vasilis Syrgkanis , Zhiwei Steven Wu

Generative models offer a direct way of modeling complex data. Energy-based models attempt to encode the statistical correlations observed in the data at the level of the Boltzmann weight associated with an energy function in the form of a…

无序系统与神经网络 · 物理学 2024-04-10 Aurélien Decelle , Cyril Furtlehner , Alfonso De Jesus Navas Gómez , Beatriz Seoane

This paper considers the control of uncertain systems that are operated under limited resource factors, such as battery life or hardware longevity. We consider here resource-aware self-triggered control techniques that schedule system…

系统与控制 · 电气工程与系统科学 2021-03-11 Yingzhao Lian , Yuning Jiang , Naomi Stricker , Lothar Thiele , Colin N. Jones

We introduce an Attention Overload Model that captures the idea that alternatives compete for the decision maker's attention, and hence the attention that each alternative receives decreases as the choice problem becomes larger. Using this…

理论经济学 · 经济学 2024-09-17 Matias D. Cattaneo , Paul Cheung , Xinwei Ma , Yusufcan Masatlioglu

Model-based reinforcement learning is an effective approach for controlling an unknown system. It is based on a longstanding pipeline familiar to the control community in which one performs experiments on the environment to collect a…

系统与控制 · 电气工程与系统科学 2024-08-14 Bruce D. Lee , Ingvar Ziemann , George J. Pappas , Nikolai Matni

We present two models of perpetrators' decision-making in extracting resources from a protected area. It is assumed that the authorities conduct surveillance to counter the extraction activities, and that perpetrators choose their…

最优化与控制 · 数学 2020-03-04 Elliot Cartee , Alexander Vladimirsky

This article proposes a fundamental methodological shift in the modelling of policy interventions for sustainability transitions in order to account for complexity (e.g. self-reinforcing mechanism arising from multi-agent interactions) and…

物理与社会 · 物理学 2016-03-23 J. -F. Mercure , H. Pollitt , A. M. Bassi , J. E Viñuales , N. R. Edwards

A multiagent system may be thought of as an artificial society of autonomous software agents and we can apply concepts borrowed from welfare economics and social choice theory to assess the social welfare of such an agent society. In this…

多智能体系统 · 计算机科学 2011-09-30 U. Endriss , N. Maudet , F. Sadri , F. Toni

We analyze the long term behavior of interacting populations which can be controlled through harvesting. The dynamics is assumed to be discrete in time and stochastic due to the effect of environmental fluctuations. We present extinction…

种群与进化 · 定量生物学 2021-02-18 Alexandru Hening

The rapid growth of wireless and mobile Internet has led to wide applications of exchanging resources over network, in which how to fairly allocate resources has become a critical challenge. To motivate sharing, a BD Mechanism is proposed…

计算机科学与博弈论 · 计算机科学 2019-05-07 Xiang Yan , Wei Zhu

In the context of quantum technologies over continuous variables, Gaussian states and operations are typically regarded as freely available, as they are relatively easily accessible experimentally. In contrast, the generation of…

量子物理 · 物理学 2022-07-13 Oliver Hahn , Patric Holmvall , Pascal Stadler , Giulia Ferrini , Alessandro Ferraro

This paper studies finite-time optimal consumption-investment problems with power, logarithmic and exponential utilities, in a regime switching market with random coefficients, subject to coupled constraints on the consumption and…

概率论 · 数学 2022-11-11 Ying Hu , Xiaomin Shi , Zuo Quan Xu

We study the policy iteration algorithm (PIA) for entropy-regularized stochastic control problems on an infinite time horizon with a large discount rate, focusing on two main scenarios. First, we analyze PIA with bounded coefficients where…

最优化与控制 · 数学 2025-05-28 Hung Vinh Tran , Zhenhua Wang , Yuming Paul Zhang

Incomplete knowledge of the environment leads an agent to make decisions under uncertainty. One of the major dilemmas in Reinforcement Learning (RL) where an autonomous agent has to balance two contrasting needs in making its decisions is:…

机器学习 · 统计学 2024-02-21 Valentina Zangirolami , Matteo Borrotti

We introduce and study a non-equilibrium continuous-time dynamical model of the price of a single asset traded by a population of heterogeneous interacting agents in the presence of uncertainty and regulatory constraints. The model takes…

适应与自组织系统 · 物理学 2009-04-23 V. I. Yukalov , D. Sornette , E. P. Yukalova

Dynamic Population Games (DPGs) provide a tractable framework for modeling strategic interactions in large populations of self-interested agents, and have been successfully applied to the design of Karma economies, a class of fair…

计算机科学与博弈论 · 计算机科学 2026-05-13 Matteo Cederle , Saverio Bolognani , Gian Antonio Susto

Sequential allocation is a simple and widely studied mechanism to allocate indivisible items in turns to agents according to a pre-specified picking sequence of agents. At each turn, the current agent in the picking sequence picks its most…

数据结构与算法 · 计算机科学 2019-09-17 Mingyu Xiao , Jiaxing Ling

Model predictive control can optimally deal with nonlinear systems under consideration of constraints. The control performance depends on the model accuracy and the prediction horizon. Recent advances propose to use reinforcement learning…

机器学习 · 计算机科学 2024-11-01 Dean Brandner , Sergio Lucia

This study rigorously investigates the Keynesian cross model of a national economy with a focus on the dynamic relationship between government spending and economic equilibrium. The model consists of two ordinary differential equations…

综合经济学 · 经济学 2023-03-21 Xinyu Li

We study merchant energy production modeled as a compound switching and timing option. The resulting Markov decision process is intractable. State-of-the-art approximate dynamic programming methods applied to realistic instances of this…

最优化与控制 · 数学 2020-01-01 Bo Yang , Selvaprabu Nadarajah , Nicola Secomandi