中文
相关论文

相关论文: Exploration noise for learning linear-quadratic me…

200 篇论文

Consistency training regularizes a model by enforcing predictions of original and perturbed inputs to be similar. Previous studies have proposed various augmentation methods for the perturbation but are limited in that they are agnostic to…

计算与语言 · 计算机科学 2022-04-29 Jungsoo Park , Gyuwan Kim , Jaewoo Kang

Non-cooperative and cooperative games with a very large number of players have many applications but remain generally intractable when the number of players increases. Introduced by Lasry and Lions, and Huang, Caines and Malham\'e, Mean…

We construct a semi-Lagrangian scheme for first-order, time-dependent, and non-local Mean Field Games. The convergence of the scheme to a weak solution of the system is analyzed by exploiting a key monotonicity property. To solve the…

数值分析 · 数学 2026-05-12 Elisabetta Carlini , Valentina Coscetti

This paper considers mean field games with optimal stopping time (OSMFGs) where agents make optimal exit decisions, the coupled obstacle and Fokker-Planck equations in such models pose challenges versus classic MFGs. This paper proposes a…

数值分析 · 数学 2023-10-10 Chengfeng Shen , Yifan Luo , Zhennan Zhou

In this paper we explore the impact of quantiles on optimal strategies under state dynamics driven by both individual noise, common noise and Poisson jumps. We first establish an optimality system satisfied the quantile process under jump…

最优化与控制 · 数学 2017-08-22 Hamidou Tembine

Here, we consider a regularized mean-field game model that features a low-order regularization. We prove the existence of solutions with positive density. To do so, we combine a priori estimates with the continuation method. In contrast…

Reinforcement learning is a powerful tool to learn the optimal policy of possibly multiple agents by interacting with the environment. As the number of agents grow to be very large, the system can be approximated by a mean-field problem.…

最优化与控制 · 数学 2020-08-18 Weichen Wang , Jiequn Han , Zhuoran Yang , Zhaoran Wang

This paper proposes using a linear function approximator, rather than a deep neural network (DNN), to bias a Monte Carlo tree search (MCTS) player for general games. This is unlikely to match the potential raw playing strength of DNNs, but…

人工智能 · 计算机科学 2019-03-22 Dennis J. N. J. Soemers , Éric Piette , Cameron Browne

Variational inference with a factorized Gaussian posterior estimate is a widely used approach for learning parameters and hidden variables. Empirically, a regularizing effect can be observed that is poorly understood. In this work, we show…

机器学习 · 计算机科学 2019-09-04 Julius Kunze , Louis Kirsch , Hippolyt Ritter , David Barber

This paper presents results from the design and testing of an educational version of Quantum Moves, a Scientific Discovery Game that allows players to help solve authentic scientific challenges in the effort to develop a quantum computer.…

物理教育 · 物理学 2015-11-06 Rikke Magnussen , Sidse Damgaard Hansen , Tilo Planke , Jacob Friis Sherson

In classical mechanics, a natural way to simplify a many-body problem is to ``replace'' some of the elements of the composite system with surrogate \textit{force fields}. In the realm of quantum mechanics, however, such a description is…

量子物理 · 物理学 2021-08-06 Piotr Szańkowski

Synergy between evolutionary dynamics of cooperation and fluctuating state of shared resource being consumed by the cooperators is essential for averting the tragedy of the commons. Not only in humans, but also in the cognitively-limited…

种群与进化 · 定量生物学 2025-12-05 Samrat Sohel Mondal , Sagar Chakraborty

In an inverse game problem, one needs to infer the cost function of the players in a game such that a desired joint strategy is a Nash equilibrium. We study the inverse game problem for a class of multiplayer matrix games, where the cost…

计算机科学与博弈论 · 计算机科学 2022-10-17 Yue Yu , Jonathan Salfity , David Fridovich-Keil , Ufuk Topcu

This paper is concerned with a class of linear-quadratic stochastic large-population problems with partial information, where the individual agent only has access to a noisy observation process related to the state. The dynamics of each…

最优化与控制 · 数学 2024-08-20 Min Li , Na Li , Zhen Wu

We present a study of a phase-separation process induced by the presence of spatially-correlated multiplicative noise. We develop a mean-field approach suitable for conserved-order-parameter systems and use it to obtain the phase diagram of…

凝聚态物理 · 物理学 2009-10-31 M. Ibanes , J. Garcia-Ojalvo , R. Toral , J. M. Sancho

We present a general computation model inspired in the notion of information hiding in software engineering. This model has the form of a game which we call quiz game. It allows in a uniform way to prove exponential lower bounds for several…

Mean-field theory has been extensively explored in decision analysis of {large-scale} (LS) systems but traditionally in ``pure" cooperative or competitive settings. This leads to the so-called mean-field game (MG) or mean-field team (MT).…

最优化与控制 · 数学 2023-06-30 Huang Jianhui , Qiu Zhenghong , Wang Shujun , Wu Zhen

Fictitious play (FP) is a well-studied algorithm that enables agents to learn Nash equilibrium in games with certain reward structures. However, when agents have no prior knowledge of the reward functions, FP faces a major challenge: the…

计算机科学与博弈论 · 计算机科学 2025-08-28 Semih Kara , Tamer Başar

Despite rapid advances in speech recognition, current models remain brittle to superficial perturbations to their inputs. Small amounts of noise can destroy the performance of an otherwise state-of-the-art model. To harden models against…

音频与语音处理 · 电气工程与系统科学 2018-07-19 Davis Liang , Zhiheng Huang , Zachary C. Lipton

Procedural noise is a fundamental component of computer graphics pipelines, offering a flexible way to generate textures that exhibit "natural" random variation. Many different types of noise exist, each produced by a separate algorithm. In…