中文
相关论文

相关论文: Rational Decisions, Random Matrices and Spin Glass…

200 篇论文

We report a novel, computationally efficient approach for solving hard nonlinear problems of reinforcement learning (RL). Here we combine umbrella sampling, from computational physics/chemistry, with optimal control methods. The approach is…

机器学习 · 计算机科学 2025-02-28 Egor E. Nuzhin , Nikolai V. Brilliantov

The paper addresses general constrained and non-linear optimization problems. For some of these notoriously hard problems, there exists a reformulation as an unconstrained, global optimization problem. We illustrate the transformation, and…

最优化与控制 · 数学 2023-06-13 Vladimir Norkin , Alois Pichler

Existing methods for nonlinear robust control often use scenario-based approaches to formulate the control problem as nonlinear optimization problems. Increasing the number of scenarios improves robustness, while increasing the size of the…

最优化与控制 · 数学 2023-06-09 Marta Zagorowska , Paola Falugi , Edward O'Dwyer , Eric C. Kerrigan

In these two lectures I review our theoretical understanding of spin glasses paying a particular attention to the basic physical ideas. We introduce the replica method and we describe its probabilistic consequences (we stress the recently…

无序系统与神经网络 · 物理学 2007-05-23 Giorgio Parisi

Coordination is a desirable feature in many multi-agent systems such as robotic and socioeconomic networks. We consider a task allocation problem as a binary networked coordination game over an undirected regular graph. Each agent in the…

系统与控制 · 电气工程与系统科学 2023-10-02 Yifei Zhang , Marcos M. Vasconcelos

A spin glass is a diluted magnetic material in which the magnetic moments are randomly interacting, with a huge number of metastable states which prevent reaching equilibrium. Spin-glass models are conceptually simple, but require very…

无序系统与神经网络 · 物理学 2023-03-03 Eric Vincent

The choice of the parameter value for regularized inverse problems is critical to the results and remains a topic of interest. This article explores a criterion for selecting a good parameter value by maximizing the probability of the data,…

数值分析 · 数学 2020-02-11 Toby Sanders , Rodrigo B. Platte , Robert D. Skeel

We study the spacial and temporal multiscale properties of complex systems. We present accelerated algorithms for dilute spin glasses and display explicitly their relation to the effective dynamics of specific collective degrees of freedom…

凝聚态物理 · 物理学 2008-02-03 N. Persky , S. Solomon

We consider the problem of strategic classification, where a learner must build a model to classify agents based on features that have been strategically modified. Previous work in this area has concentrated on the case when the learner is…

机器学习 · 计算机科学 2025-05-19 Jack Geary , Henry Gouk

This paper provides a non-robust interpretation of the distributionally robust optimization (DRO) problem by relating the distributional uncertainties to the chance probabilities. Our analysis allows a decision-maker to interpret the size…

最优化与控制 · 数学 2020-09-22 Qi Wu , Shumin Ma , Cheuk Hang Leung , Wei Liu , Nanbo Peng

The assortment problem in revenue management is the problem of deciding which subset of products to offer to consumers in order to maximise revenue. A simple and natural strategy is to select the best assortment out of all those that are…

数据结构与算法 · 计算机科学 2019-02-22 Gerardo Berbeglia , Gwenaël Joret

In this paper we consider multiple constrained resource allocation problems, where the constraints can be specified by formulating activity dependency restrictions or by using game-theoretic models. All the problems are focused on generic…

数据结构与算法 · 计算机科学 2009-06-19 Mugurel Ionut Andreica , Madalina Ecaterina Andreica , Costel Visan

Research in reinforcement learning has produced algorithms for optimal decision making under uncertainty that fall within two main types. The first employs a Bayesian framework, where optimality improves with increased computational time.…

机器学习 · 统计学 2011-09-22 Christos Dimitrakakis

Motion planning under differential constraints is a classic problem in robotics. To date, the state of the art is represented by sampling-based techniques, with the Rapidly-exploring Random Tree algorithm as a leading example. Yet, the…

机器人学 · 计算机科学 2015-03-03 Edward Schmerling , Lucas Janson , Marco Pavone

We consider reinforcement learning in changing Markov Decision Processes where both the state-transition probabilities and the reward functions may vary over time. For this problem setting, we propose an algorithm using a sliding window…

机器学习 · 计算机科学 2018-05-28 Pratik Gajane , Ronald Ortner , Peter Auer

We consider the problem of globally minimizing the sum of many rational functions over a given compact semialgebraic set. The number of terms can be large (10 to 100), the degree of each term should be small (up to 10), and the number of…

最优化与控制 · 数学 2011-02-25 Florian Bugarin , Didier Henrion , Jean-Bernard Lasserre

This paper shows how we can combine logical representations of actions and decision theory in such a manner that seems natural for both. In particular we assume an axiomatization of the domain in terms of situation calculus, using what is…

人工智能 · 计算机科学 2013-02-18 David L. Poole

Robust optimization(RO) is an important tool for handling optimization problem with uncertainty. The main objective of RO is to solve optimization problems due to uncertainty associated with constraints satisfying all realizations of…

最优化与控制 · 数学 2025-04-02 Parthasarathi Mondal , Akshay Kumar Ojha

Strategic Decision-Making is always challenging because it is inherently uncertain, ambiguous, risky, and complex. It is the art of possibility. We develop a systematic taxonomy of decision-making frames that consists of 6 bases, 18…

人工智能 · 计算机科学 2022-10-25 Caesar Wu , Kotagiri Ramamohanarao , Rui Zhang , Pascal Bouvry

In this article we introduce the use of recently developed min/max-plus techniques in order to solve the optimal attitude estimation problem in filtering for nonlinear systems on the special orthogonal (SO(3)) group. This work helps obtain…

最优化与控制 · 数学 2012-11-08 Srinivas Sridharan , William M. McEneaney