中文
相关论文

相关论文: Learning the Sherrington-Kirkpatrick Model Even at…

200 篇论文

We consider online learning for minimizing regret in unknown, episodic Markov decision processes (MDPs) with continuous states and actions. We develop variants of the UCRL and posterior sampling algorithms that employ nonparametric Gaussian…

机器学习 · 计算机科学 2019-01-04 Sayak Ray Chowdhury , Aditya Gopalan

Unsupervised machine learning methods are used to identify structural changes using the melting point transition in classical molecular dynamics simulations as an example application of the approach. Dimensionality reduction and clustering…

计算物理 · 物理学 2018-12-06 Nicholas Walker , Ka-Ming Tam , Brian Novak , M. Jarrell

The theory of learning under the uniform distribution is rich and deep, with connections to cryptography, computational complexity, and the analysis of boolean functions to name a few areas. This theory however is very limited due to the…

机器学习 · 计算机科学 2015-06-15 Varun Kanade , Elchanan Mossel

In this paper, we study an online regularized learning algorithm in a reproducing kernel Hilbert spaces (RKHS) based on a class of dependent processes. We choose such a process where the degree of dependence is measured by mixing…

机器学习 · 统计学 2025-07-09 Priyanka Roy , Susanne Saminger-Platz

Complex numerical weather prediction models incorporate a variety of physical processes, each described by multiple alternative physical schemes with specific parameters. The selection of the physical schemes and the choice of the…

数值分析 · 计算机科学 2018-02-23 Azam Moosavi , Vishwas Rao , Adrian Sandu

We develop a Machine Learning Inversion method for analyzing scattering functions of mechanically driven polymers and extracting the corresponding feature parameters, which include energy parameters and conformation variables. The polymer…

软凝聚态物质 · 物理学 2025-11-21 Lijie Ding , Chi-Huan Tung , Bobby G. Sumpter , Wei-Ren Chen , Changwoo Do

Training reinforcement learning (RL) agents using scalar reward signals is often infeasible when an environment has sparse and non-Markovian rewards. Moreover, handcrafting these reward functions before training is prone to…

机器学习 · 计算机科学 2023-10-04 Alessandro Abate , Yousif Almulla , James Fox , David Hyland , Michael Wooldridge

We consider un-discounted reinforcement learning (RL) in Markov decision processes (MDPs) under drifting non-stationarity, i.e., both the reward and state transition distributions are allowed to evolve over time, as long as their respective…

机器学习 · 计算机科学 2020-06-26 Wang Chi Cheung , David Simchi-Levi , Ruihao Zhu

In this paper, we study reinforcement learning in Markov Decision Processes with Probabilistic Reward Machines (PRMs), a form of non-Markovian reward commonly found in robotics tasks. We design an algorithm for PRMs that achieves a regret…

机器学习 · 统计学 2024-08-21 Xiaofeng Lin , Xuezhou Zhang

Accurately determining the underlying physical parameters of individual elements in integrated photonics is increasingly difficult as device architectures become more complex. Inferring these parameters directly from spectral measurements…

We consider the task of Inverse Reinforcement Learning in Contextual Markov Decision Processes (MDPs). In this setting, contexts, which define the reward and transition kernel, are sampled from a distribution. In addition, although the…

机器学习 · 计算机科学 2021-01-01 Stav Belogolovsky , Philip Korsunsky , Shie Mannor , Chen Tessler , Tom Zahavy

We describe how the couplings in an asynchronous kinetic Ising model can be inferred. We consider two cases, one in which we know both the spin history and the update times and one in which we only know the spin history. For the first case,…

数据分析、统计与概率 · 物理学 2015-06-11 Hong-Li Zeng , Mikko Alava , Erik Aurell , John Hertz , Yasser Roudi

We present new MCMC algorithms for computing the posterior distributions and expectations of the unknown variables in undirected graphical models with regular structure. For demonstration purposes, we focus on Markov Random Fields (MRFs).…

统计计算 · 统计学 2012-07-19 Firas Hamze , Nando de Freitas

A major component of overfitting in model-free reinforcement learning (RL) involves the case where the agent may mistakenly correlate reward with certain spurious features from the observations generated by the Markov Decision Process…

机器学习 · 计算机科学 2020-01-01 Xingyou Song , Yiding Jiang , Stephen Tu , Yilun Du , Behnam Neyshabur

We study the mixing time of systematic scan Glauber dynamics Ising model on the complete graph. On the complete graph $K_n$, at each time, $k \leq n$ vertices are chosen uniformly random and are updated one by one according to the uniformly…

概率论 · 数学 2024-11-11 Sanghak Jeon

The Ising model on a $restricted$ scale-free network (SFN) has been studied employing Monte Carlo simulations. This network is described by a power-law degree distribution in the form $P(k)~k^{-\alpha}$, and is called restricted, because…

统计力学 · 物理学 2023-05-24 R. A. Dumer , M. Godoy

We develop a method to learn physical systems from data that employs feedforward neural networks and whose predictions comply with the first and second principles of thermodynamics. The method employs a minimum amount of data by enforcing…

机器学习 · 计算机科学 2020-11-16 Quercus Hernández , Alberto Badias , David Gonzalez , Francisco Chinesta , Elias Cueto

We study off-policy reinforcement learning for controlling continuous-time Markov diffusion processes with discrete-time observations and actions. We consider model-free algorithms with function approximation that learn value and advantage…

机器学习 · 计算机科学 2026-04-17 Wenlong Mou

Hamiltonian parameter estimation is crucial in condensed matter physics, but time and cost consuming in terms of resources used. With advances in observation techniques, high-resolution images with more detailed information are obtained,…

无序系统与神经网络 · 物理学 2019-11-15 Dingchen Wang , Songrui Wei , Anran Yuan , Fanghua Tian , Kaiyan Cao , Qizhong Zhao , Dezhen Xue , Sen Yang

In this paper, we propose a learning-based Model Predictive Control (MPC) approach for the polytopic Linear Parameter-Varying (LPV) systems with inexact scheduling parameters (as exogenous signals with inexact bounds), where the Linear Time…

系统与控制 · 电气工程与系统科学 2022-06-13 Hossein Nejatbakhsh Esfahani , Sebastien Gros