中文
相关论文

相关论文: Learning the Sherrington-Kirkpatrick Model Even at…

200 篇论文

Boltzmann Machines (BMs) are graphical models with interconnected binary units, employed for the unsupervised modeling of data distributions. When trained on real data, BMs show the tendency to behave like critical systems, displaying a…

无序系统与神经网络 · 物理学 2024-06-28 Enrico Ventura , Simona Cocco , Rémi Monasson , Francesco Zamponi

Projecting climate change is a generalization problem: we extrapolate the recent past using physical models across past, present, and future climates. Current climate models require representations of processes that occur at scales smaller…

We consider the problem of learning by demonstration from agents acting in unknown stochastic Markov environments or games. Our aim is to estimate agent preferences in order to construct improved policies for the same task that the agents…

机器学习 · 计算机科学 2014-08-12 Aristide Tossou , Christos Dimitrakakis

We consider the problem of learning by demonstration from agents acting in unknown stochastic Markov environments or games. Our aim is to estimate agent preferences in order to construct improved policies for the same task that the agents…

机器学习 · 统计学 2013-07-16 Aristide C. Y. Tossou , Christos Dimitrakakis

We demonstrate the applicability of the $\epsilon$-convergence algorithm in extracting the critical temperatures and critical exponents of three-dimensional Ising models. We analyze the low temperature magnetization as well as high…

统计力学 · 物理学 2024-10-22 M V Vismaya , M V Sangaranarayanan

Inverse reinforcement learning (IRL) infers a reward function from demonstrations, allowing for policy improvement and generalization. However, despite much recent interest in IRL, little work has been done to understand the minimum set of…

机器学习 · 计算机科学 2019-08-19 Daniel S. Brown , Scott Niekum

We study the model-based reward-free reinforcement learning with linear function approximation for episodic Markov decision processes (MDPs). In this setting, the agent works in two phases. In the exploration phase, the agent interacts with…

机器学习 · 计算机科学 2022-01-03 Weitong Zhang , Dongruo Zhou , Quanquan Gu

Simulated tempering is popular method of allowing MCMC algorithms to move between modes of a multimodal target density {\pi}. One problem with simulated tempering for multimodal targets is that the weights of the various modes change for…

统计计算 · 统计学 2019-02-12 Nicholas G. Tawn , Gareth O. Roberts , Jeffrey S. Rosenthal

In this paper a multi-scale version of the Sherrington and Kirkpatrick model is introduced and studied. The pressure per particle in the thermodynamical limit is proved to obey a variational principle of Parisi type. The result is achieved…

数学物理 · 物理学 2019-02-20 Pierluigi Contucci , Emanuele Mingione

In real world settings, numerous constraints are present which are hard to specify mathematically. However, for the real world deployment of reinforcement learning (RL), it is critical that RL agents are aware of these constraints, so that…

机器学习 · 计算机科学 2021-05-24 Usman Anwar , Shehryar Malik , Alireza Aghasi , Ali Ahmed

Machine learning has become a central technique for modeling in science and engineering, either complementing or as surrogates to physics-based models. Significant efforts have recently been devoted to models capable of predicting field…

统计力学 · 物理学 2024-12-06 Brian H. Lee , Kat Nykiel , Ava E. Hallberg , Brice Rider , Alejandro Strachan

We propose a simple algorithm to train stochastic neural networks to draw samples from given target distributions for probabilistic inference. Our method is based on iteratively adjusting the neural network parameters so that the output…

机器学习 · 统计学 2016-11-29 Dilin Wang , Qiang Liu

Single crystal inelastic neutron scattering data contain rich information about the structure and dynamics of a material. Yet the challenge of matching sophisticated theoretical models with large data volumes is compounded by computational…

We introduce a novel method that enables parameter-efficient transfer and multi-task learning with deep neural networks. The basic approach is to learn a model patch - a small set of parameters - that will specialize to each task, instead…

机器学习 · 计算机科学 2019-02-26 Pramod Kaushik Mudrakarta , Mark Sandler , Andrey Zhmoginov , Andrew Howard

We show that the performance of critical quantum metrology protocols, counter-intuitively, can be enhanced by finite temperature. We consider a toy-model squeezing Hamiltonian, the Lipkin-Meshkov-Glick model and the paradigmatic Ising…

量子物理 · 物理学 2024-02-23 Laurin Ostermann , Karol Gietka

We apply various unsupervised machine learning methods for phase classification to investigate the finite-temperature phase diagram of the spinless Falicov-Kimball model in two dimensions. Using only particle occupation snapshots from Monte…

强关联电子 · 物理学 2025-05-27 Lukáš Frk , Pavel Baláž , Elguja Archemashvili , Martin Žonda

We study inverse reinforcement learning (IRL) and imitation learning (IM), the problems of recovering a reward or policy function from expert's demonstrated trajectories. We propose a new way to improve the learning process by adding a…

机器学习 · 计算机科学 2022-08-23 The Viet Bui , Tien Mai , Patrick Jaillet

We study the Sherrington--Kirkpatrick model, both above and below the De Almeida Thouless line, by using a modified version of the Parallel Tempering algorithm in which the system is allowed to move between different values of the magnetic…

统计力学 · 物理学 2009-11-07 Alain Billoire , Barbara Coluzzi

We consider a reinforcement learning (RL) setting in which the agent interacts with a sequence of episodic MDPs. At the start of each episode the agent has access to some side-information or context that determines the dynamics of the MDP…

机器学习 · 统计学 2019-10-24 Aditya Modi , Nan Jiang , Satinder Singh , Ambuj Tewari

Incremental methods for structure learning of pairwise Markov random fields (MRFs), such as grafting, improve scalability by avoiding inference over the entire feature space in each optimization step. Instead, inference is performed over an…

机器学习 · 计算机科学 2018-05-22 Walid Chaabene , Bert Huang