中文
相关论文

相关论文: Representation of Reinforcement Learning Policies …

200 篇论文

We present a general framework to learn functions in tensor product reproducing kernel Hilbert spaces (TP-RKHSs). The methodology is based on a novel representer theorem suitable for existing as well as new spectral penalties for tensors.…

机器学习 · 计算机科学 2013-10-21 Marco Signoretto , Lieven De Lathauwer , Johan A. K. Suykens

A Hilbert space embedding of a distribution---in short, a kernel mean embedding---has recently emerged as a powerful tool for machine learning and inference. The basic idea behind this framework is to map distributions into a reproducing…

机器学习 · 统计学 2020-12-15 Krikamol Muandet , Kenji Fukumizu , Bharath Sriperumbudur , Bernhard Schölkopf

In this paper, we consider the reproducing property in Reproducing Kernel Hilbert Spaces (RKHS). We establish a reproducing property for the closure of the class of combinations of composition operators under minimal conditions. This allows…

统计理论 · 数学 2025-04-01 Fatima-Zahrae El-Boukkouri , Josselin Garnier , Olivier Roustant

We review machine learning methods employing positive definite kernels. These methods formulate learning and estimation problems in a reproducing kernel Hilbert space (RKHS) of functions defined on the data domain, expanded in terms of a…

统计理论 · 数学 2009-09-29 Thomas Hofmann , Bernhard Schölkopf , Alexander J. Smola

This paper studies convergence rates for some value function approximations that arise in a collection of reproducing kernel Hilbert spaces (RKHS) $H(\Omega)$. By casting an optimal control problem in a specific class of native spaces,…

系统与控制 · 电气工程与系统科学 2023-11-20 Ali Bouland , Shengyuan Niu , Sai Tej Paruchuri , Andrew Kurdila , John Burns , Eugenio Schuster

A nonparametric approach for policy learning for POMDPs is proposed. The approach represents distributions over the states, observations, and actions as embeddings in feature spaces, which are reproducing kernel Hilbert spaces.…

机器学习 · 计算机科学 2012-10-19 Yu Nishiyama , Abdeslam Boularias , Arthur Gretton , Kenji Fukumizu

In reinforcement learning, we encode the potential behaviors of an agent interacting with an environment into an infinite set of policies, the policy space, typically represented by a family of parametric functions. Dealing with such a…

机器学习 · 计算机科学 2022-02-23 Mirco Mutti , Stefano Del Col , Marcello Restelli

Reinforcement learning (RL) has demonstrated its ability to solve high dimensional tasks by leveraging non-linear function approximators. However, these successes are mostly achieved by 'black-box' policies in simulated domains. When…

机器学习 · 计算机科学 2021-11-19 Riad Akrour , Davide Tateo , Jan Peters

Since its introduction, the Discrete Variable Representation (DVR) basis set has become an invaluable representation of state vectors and Hermitian operators in non-relativistic quantum dynamics and spectroscopy calculations. On the other…

计算物理 · 物理学 2014-05-30 Hamse Mussa

This paper aims at the algorithmic/theoretical core of reinforcement learning (RL) by introducing the novel class of proximal Bellman mappings. These mappings are defined in reproducing kernel Hilbert spaces (RKHSs), to benefit from the…

信号处理 · 电气工程与系统科学 2023-09-15 Yuki Akiyama , Konstantinos Slavakis

This short technical report presents some learning theory results on vector-valued reproducing kernel Hilbert space (RKHS) regression, where the input space is allowed to be non-compact and the output space is a (possibly…

机器学习 · 统计学 2022-02-17 Junhyunng Park , Krikamol Muandet

Traditional machine learning models, particularly neural networks, are rooted in finite-dimensional parameter spaces and nonlinear function approximations. This report explores an alternative formulation where learning tasks are expressed…

机器学习 · 计算机科学 2025-07-30 Andrew Kiruluta , Andreas Lemos , Priscilla Burity

Reproducing kernel Hilbert spaces (RKHSs) are special Hilbert spaces where all the evaluation functionals are linear and bounded. They are in one-to-one correspondence with positive definite maps called kernels. Stable RKHSs enjoy the…

系统与控制 · 电气工程与系统科学 2023-05-04 Mauro Bisiacco , Gianluigi Pillonetto

A central challenge in reinforcement learning (RL) is to learn models that generalize beyond the tasks on which they are trained, a goal traditionally pursued through multi-task and meta RL. Recently, transformer architectures have emerged…

机器学习 · 计算机科学 2026-05-12 Bowen He , Juncheng Dong , Lin Lin , Xiang Cheng

Motivated by the success of reinforcement learning (RL) for discrete-time tasks such as AlphaGo and Atari games, there has been a recent surge of interest in using RL for continuous-time control of physical systems (cf. many challenging…

最优化与控制 · 数学 2018-12-03 Motoya Ohnishi , Masahiro Yukawa , Mikael Johansson , Masashi Sugiyama

This paper develops a frequentist solution to the functional calibration problem, where the value of a calibration parameter in a computer model is allowed to vary with the value of control variables in the physical system. The need of…

统计方法学 · 统计学 2021-07-20 Rui Tuo , Shiyuan He , Arash Pourhabib , Yu Ding , Jianhua Z. Huang

Reduced modeling of a computationally demanding dynamical system aims at approximating its trajectories, while optimizing the trade-off between accuracy and computational complexity. In this work, we propose to achieve such an approximation…

机器学习 · 统计学 2025-02-20 Patrick Héas , Cédric Herzet , Benoit Combès

Predictive State Representations (PSRs) are an expressive class of models for controlled stochastic processes. PSRs represent state as a set of predictions of future observable events. Because PSRs are defined entirely in terms of…

机器学习 · 计算机科学 2013-09-27 Byron Boots , Geoffrey Gordon , Arthur Gretton

Supervised learning in reproducing kernel Hilbert space (RKHS) and vector-valued RKHS (vvRKHS) has been investigated for more than 30 years. In this paper, we provide a new twist to this rich literature by generalizing supervised learning…

机器学习 · 统计学 2024-06-27 Yuka Hashimoto , Masahiro Ikeda , Hachem Kadri

Reinforcement learning (RL) has shown empirical success in various real world settings with complex models and large state-action spaces. The existing analytical results, however, typically focus on settings with a small number of…

机器学习 · 计算机科学 2024-03-15 Sattar Vakili , Julia Olkhovskaya