中文
相关论文

相关论文: An Approximation of the First Order Marcum $Q$-Fun…

200 篇论文

First-order Marcum $Q$-function is observed in various problem formulations. However, it is not an easy-to-handle function. For this reason, in this paper, we first present a semi-linear approximation of the Marcum $Q$-function. Our…

信息论 · 计算机科学 2020-12-11 Hao Guo , Behrooz Makki , Mohamed-Slim Alouini , Tommy Svensson

We propose an approach to lifted approximate inference for first-order probabilistic models, such as Markov logic networks. It is based on performing exact lifted inference in a simplified first-order model, which is found by relaxing…

人工智能 · 计算机科学 2012-10-19 Guy Van den Broeck , Arthur Choi , Adnan Darwiche

$Q$-learning with function approximation is one of the most empirically successful while theoretically mysterious reinforcement learning (RL) algorithms, and was identified in Sutton (1999) as one of the most important theoretical open…

机器学习 · 计算机科学 2022-05-04 Zaiwei Chen , John Paul Clarke , Siva Theja Maguluri

The paper introduces the first formulation of convex Q-learning for Markov decision processes with function approximation. The algorithms and theory rest on a relaxation of a dual of Manne's celebrated linear programming characterization of…

最优化与控制 · 数学 2023-09-12 Fan Lu , Sean Meyn

This work presents analytic solutions for a useful integral in wireless communications, which involves the Marcum $Q{-}$function in combination with an exponential function and arbitrary power terms. The derived expressions have a rather…

In this paper we study the right differentiability of a parametric infimum function over a parametric set defined by equality constraints. We present a new theorem with sufficient conditions for the right differentiability with respect to…

最优化与控制 · 数学 2023-06-22 Kevin Sturm

Approximate inference in dynamic systems is the problem of estimating the state of the system given a sequence of actions and partial observations. High precision estimation is fundamental in many applications like diagnosis, natural…

人工智能 · 计算机科学 2012-06-18 Hannaneh Hajishirzi , Eyal Amir

Matrix rank minimization problems are gaining a plenty of recent attention in both mathematical and engineering fields. This class of problems, arising in various and across-discipline applications, is known to be NP-hard in general. In…

最优化与控制 · 数学 2010-10-06 Yun-Bin Zhao

In this paper, we present a comprehensive study of the monotonicity and log-concavity of the generalized Marcum and Nuttall Q-functions. More precisely, a simple probabilistic method is firstly given to prove the monotonicity of these two…

信息论 · 计算机科学 2015-03-13 Yin Sun , Arpad Baricz , Shidong Zhou

Methods and an algorithm for computing the generalized Marcum $Q-$function ($Q_{\mu}(x,y)$) and the complementary function ($P_{\mu}(x,y)$) are described. These functions appear in problems of different technical and scientific areas such…

数学软件 · 计算机科学 2013-11-05 A. Gil , J. Segura , N. M. Temme

Novel analytic solutions are derived for integrals that involve the generalized Marcum Q-function, exponential functions and arbitrary powers. Simple closed-form expressions are also derived for the specific cases of the generic integrals.…

信息论 · 计算机科学 2023-07-19 Paschalis C. Sofotasios , Sami Muhaidat , George K. Karagiannidis , Bayan S. Sharif

In this study, we consider the application of max-plus-linear approximators for Q-function in offline reinforcement learning of discounted Markov decision processes. In particular, we incorporate these approximators to propose novel fitted…

最优化与控制 · 数学 2025-03-11 Y. Liu , M. A. S. Kolarijani

We introduce a new approximate solution technique for first-order Markov decision processes (FOMDPs). Representing the value function linearly w.r.t. a set of first-order basis functions, we compute suitable weights by casting the…

人工智能 · 计算机科学 2012-07-09 Scott Sanner , Craig Boutilier

We establish a continuous-time framework for analyzing Deep Q-Networks (DQNs) via stochastic control and Forward-Backward Stochastic Differential Equations (FBSDEs). Considering a continuous-time Markov Decision Process (MDP) driven by a…

机器学习 · 计算机科学 2025-05-06 Qian Qi

The growing amount of applications that generate vast amount of data in short time scales render the problem of partial monitoring, coupled with prediction, a rather fundamental one. We study the aforementioned canonical problem under the…

数据结构与算法 · 计算机科学 2016-08-02 Michalis Kallitsis , Stilian Stoev , George Michailidis

In this paper, we provide a novel algorithm for solving planning and learning problems of Markov decision processes. The proposed algorithm follows a policy iteration-type update by using a rank-one approximation of the transition…

Q-learning with neural network function approximation (neural Q-learning for short) is among the most prevalent deep reinforcement learning algorithms. Despite its empirical success, the non-asymptotic convergence rate of neural Q-learning…

机器学习 · 计算机科学 2020-03-05 Pan Xu , Quanquan Gu

We propose Q-learning with Adjoint Matching (QAM), a novel TD-based reinforcement learning (RL) algorithm that tackles a long-standing challenge in continuous-action RL: efficient optimization of an expressive diffusion or flow-matching…

机器学习 · 计算机科学 2026-05-20 Qiyang Li , Sergey Levine

We consider a class of statistical estimation problems in which we are given a random data matrix ${\boldsymbol X}\in {\mathbb R}^{n\times d}$ (and possibly some labels ${\boldsymbol y}\in{\mathbb R}^n$) and would like to estimate a…

统计计算 · 统计学 2022-01-14 Andrea Montanari , Yuchen Wu

This paper proposes a unique optimization approach for estimating the minimax rational approximation and its application for evaluating matrix functions. Our method enables the extension to generalized rational approximations and has the…

数值分析 · 数学 2025-04-03 Nir Sharon , Vinesha Peiris , Nadia Sukhorukova , Julien Ugon
‹ 上一页 1 2 3 10 下一页 ›