中文
相关论文

相关论文: The infinite Viterbi alignment and decay-convexity

200 篇论文

We define angles from-to and between infinite dimensional subspaces of a Hilbert space, inspired by the work of E. J. Hannan, 1961/1962 for general canonical correlations of stochastic processes. The spectral theory of selfadjoint operators…

数值分析 · 数学 2010-07-02 Andrew Knyazev , Abram Jujunashvili , Merico Argentati

We obtain new bounds for the optimal matching cost for empirical measures with unbounded support. For a large class of radially symmetric and rapidly decaying probability laws, we prove for the first time the asymptotic rate of convergence…

概率论 · 数学 2024-07-10 Emanuele Caglioti , Michael Goldman , Francesca Pieroni , Dario Trevisan

Optimal Dirichlet boundary control for a fractional/normal evolution with a final observation is considered. The unique existence of the solution and the first-order optimality condition of the optimal control problem are derived. The…

数值分析 · 数学 2020-07-20 Qin Zhou , Binjie Li

Many statistical $M$-estimators are based on convex optimization problems formed by the combination of a data-dependent loss function with a norm-based regularizer. We analyze the convergence rates of projected gradient and composite…

机器学习 · 统计学 2012-07-26 Alekh Agarwal , Sahand N. Negahban , Martin J. Wainwright

This paper proposes a computationally tractable algorithm for learning infinite-horizon average-reward linear mixture Markov decision processes (MDPs) under the Bellman optimality condition. Our algorithm for linear mixture MDPs achieves a…

机器学习 · 计算机科学 2024-10-22 Woojin Chae , Kihyuk Hong , Yufan Zhang , Ambuj Tewari , Dabeen Lee

In the literature, the matchings between spacetimes have been most of the times implicitly assumed to preserve some of the symmetries of the problem involved. But no definition for this kind of matching was given until recently. Loosely…

广义相对论与量子宇宙学 · 物理学 2009-11-07 Raul Vera

Motivated by Ridgway's proof of the perceptron algorithm, we study a simple subgradient method for convex inequality systems in Hilbert space. Assuming strict feasibility and bounded subgradients, we establish finite termination for several…

最优化与控制 · 数学 2026-04-27 Heinz H. Bauschke , Tran Thanh Tung

The purpose of this paper is to propose and analyze a multi-step iterative algorithm to solve a convex optimization problem and a fixed point problem posed on a Hadamard space. The convergence properties of the proposed algorithm are…

泛函分析 · 数学 2018-02-28 Muhammad Aqeel Ahmad Khan , Hafiza Arham Maqbool

We consider the nonlinear Kolmogorov equation posed in a Hilbert space $H$, not necessarily of finite dimension. This model was recently studied by Cox et al. [24] in the framework of weak convergence rates of stochastic wave models. Here,…

概率论 · 数学 2022-07-06 Javier Castro

These notes present preliminary results regarding two different approximations of linear infinite-horizon optimal control problems arising in model predictive control. Input and state trajectories are parametrized with basis functions and a…

最优化与控制 · 数学 2016-09-04 Michael Muehlebach , Raffaello D'Andrea

Latent variable models for ordinal data represent a useful tool in different fields of research in which the constructs of interest are not directly observable. In such models, problems related to the integration of the likelihood function…

统计方法学 · 统计学 2012-06-26 Silvia Bianconcini , Silvia Cagnone

Motivated by optimization with differential equations, we consider optimization problems with Hilbert spaces as decision spaces. As a consequence of their infinite dimensionality, the numerical solution necessitates finite dimensional…

最优化与控制 · 数学 2025-07-01 Danlin Li , Johannes Milz

We propose and analyze a posteriori error estimators for an optimal control problem that involves an elliptic partial differential equation as state equation and a control variable that enters the state equation as a coefficient; pointwise…

最优化与控制 · 数学 2022-03-31 Francisco Fuica , Enrique Otarola

Proximal policy optimization and trust region policy optimization (PPO and TRPO) with actor and critic parametrized by neural networks achieve significant empirical success in deep reinforcement learning. However, due to nonconvexity, the…

机器学习 · 计算机科学 2023-03-01 Boyi Liu , Qi Cai , Zhuoran Yang , Zhaoran Wang

In this paper, we study the gradient descent-ascent method for convex-concave saddle-point problems. We derive a new non-asymptotic global convergence rate in terms of distance to the solution set by using the semidefinite programming…

最优化与控制 · 数学 2022-09-19 Moslem Zamani , Hadi Abbaszadehpeivasti , Etienne de Klerk

The closed-loop stability and infinite-horizon performance of receding-horizon approximations are studied for non-stationary linear-quadratic regulator (LQR) problems. The approach is based on a lifted reformulation of the optimal control…

系统与控制 · 电气工程与系统科学 2023-09-06 Jintao Sun , Michael Cantoni

We present a new algorithm based on posterior sampling for learning in constrained Markov decision processes (CMDP) in the infinite-horizon undiscounted setting. The algorithm achieves near-optimal regret bounds while being advantageous…

机器学习 · 计算机科学 2023-09-28 Danil Provodin , Pratik Gajane , Mykola Pechenizkiy , Maurits Kaptein

We investigate the techniques and ideas used in the convergence analysis of two proximal ADMM algorithms for solving convex optimization problems involving compositions with linear operators. Besides this, we formulate a variant of the ADMM…

最优化与控制 · 数学 2019-12-20 Sebastian Banert , Radu Ioan Bot , Ernö Robert Csetnek

A deep equilibrium model (DEQ) is implicitly defined through an equilibrium point of an infinite-depth weight-tied model with an input-injection. Instead of infinite computations, it solves an equilibrium point directly with root-finding…

机器学习 · 计算机科学 2023-03-30 Zenan Ling , Xingyu Xie , Qiuhao Wang , Zongpeng Zhang , Zhouchen Lin

Relative pose estimation is fundamental for SLAM, visual localization, and 3D reconstruction. Existing Relative Pose Regression (RPR) methods face a key trade-off: feature-matching pipelines achieve high accuracy but block gradient flow via…

计算机视觉与模式识别 · 计算机科学 2026-03-23 Jun Wang , Xiaoyan Huang