English
Related papers

Related papers: Horseshoe Priors and MDP

200 papers

Bayesian fused lasso is one of the sparse Bayesian methods, which shrinks both regression coefficients and their successive differences simultaneously. In this paper, we propose a Bayesian fused lasso modeling via horseshoe prior. By…

Methodology · Statistics 2022-01-21 Yuko Kakikawa , Kaito Shimamura , Shuichi Kawano

We study robust Markov decision processes (RMDPs) with general policy parameterization under s-rectangular and non-rectangular uncertainty sets. Prior work is largely limited to tabular policies, and hence either lacks sample complexity…

Machine Learning · Computer Science 2026-02-13 Anirudh Satheesh , Ziyi Chen , Furong Huang , Heng Huang

This paper addresses the challenge of solving Constrained Markov Decision Processes (CMDPs) with $d > 1$ constraints when the transition dynamics are unknown, but samples can be drawn from a generative model. We propose a model-based…

Machine Learning · Computer Science 2025-03-11 Max Buckley , Konstantinos Papathanasiou , Andreas Spanopoulos

Priors with non-smooth log-densities, such as the l1-prior, are widely used in Bayesian inverse problems for their sparsity-inducing properties. Existing Langevin-based sampling methods typically rely on proximal mappings or smooth…

Numerical Analysis · Mathematics 2026-05-05 Ivan Cheltsov , Federico Cornalba , Clarice Poon , Tony Shardlow

We study the $(\varepsilon, \delta)$-PAC policy identification problem in finite-horizon episodic Markov Decision Processes. Existing approaches provide finite-time guarantees for approximate settings ($\varepsilon>0$) but suffer from high…

Machine Learning · Computer Science 2026-05-06 Cyrille Kone , Kevin Jamieson

The pseudo-marginal algorithm is a variant of the Metropolis--Hastings algorithm which samples asymptotically from a probability distribution when it is only possible to estimate unbiasedly an unnormalized version of its density.…

Computation · Statistics 2019-12-04 Sebastian M. Schmon , George Deligiannidis , Arnaud Doucet , Michael K. Pitt

We study infinite-horizon average-reward Markov decision processes (AMDPs) in the context of general function approximation. Specifically, we propose a novel algorithmic framework named Local-fitted Optimization with OPtimism (LOOP), which…

Machine Learning · Computer Science 2024-04-22 Jianliang He , Han Zhong , Zhuoran Yang

We present non-asymptotic two-sided bounds to the log-marginal likelihood in Bayesian inference. The classical Laplace approximation is recovered as the leading term. Our derivation permits model misspecification and allows the parameter…

Statistics Theory · Mathematics 2020-06-23 Anirban Bhattacharya , Debdeep Pati

In this study, we derive Probably Approximately Correct (PAC) bounds on the asymptotic sample-complexity for RL within the infinite-horizon Markov Decision Process (MDP) setting that are sharper than those in existing literature. The…

Machine Learning · Computer Science 2025-07-17 Mohit Prashant , Arvind Easwaran

Besov priors are nonparametric priors that can model spatially inhomogeneous functions. They are routinely used in inverse problems and imaging, where they exhibit attractive sparsity-promoting and edge-preserving features. A recent line of…

Statistics Theory · Mathematics 2023-09-11 Matteo Giordano

In Bayesian statistics, horseshoe prior has attracted increasing attention as an approach to the sparse estimation. The estimation accuracy of compressed sensing with the horseshoe prior is evaluated by statistical mechanical method. It is…

Disordered Systems and Neural Networks · Physics 2023-03-29 Yasushi Nagano , Koji Hukushima

We study the rate of Bayesian consistency for hierarchical priors consisting of prior weights on a model index set and a prior on a density model for each choice of model index. Ghosal, Lember and Van der Vaart [2] have obtained general…

Statistics Theory · Mathematics 2008-09-23 Yang Xing

We consider sparse Bayesian estimation in the classical multivariate linear regression model with $p$ regressors and $q$ response variables. In univariate Bayesian linear regression with a single response $y$, shrinkage priors which can be…

Methodology · Statistics 2018-05-21 Ray Bai , Malay Ghosh

Multiple-environment Markov decision processes (MEMDPs) equip an MDP with several probabilistic transition functions (one per possible environment) so that the state is observable but the environment is not. Previous work studies two…

Logic in Computer Science · Computer Science 2026-02-12 Benjamin Bordais , Jean-François Raskin

We propose a new policy gradient method, named homotopic policy mirror descent (HPMD), for solving discounted, infinite horizon MDPs with finite state and action spaces. HPMD performs a mirror descent type policy update with an additional…

Machine Learning · Computer Science 2022-11-30 Yan Li , Guanghui Lan , Tuo Zhao

Peskin's Immersed Boundary (IB) model and method are among one of the most important modeling tools and numerical methods. The IB method has been known to be first order accurate in the velocity. However, almost no rigorous theoretical…

Numerical Analysis · Mathematics 2022-06-07 Zhilin Li , Kejia Pan , Juan Ruiz-Álvarez

Consider the $3$-d primitive equations in a layer domain $\Omega=G \times (-h,0)$, $G=(0,1)^2$, subject to mixed Dirichlet and Neumann boundary conditions at $z=-h$ and $z=0$, respectively, and the periodic lateral boundary condition. It is…

Analysis of PDEs · Mathematics 2021-03-29 Yoshikazu Giga , Mathis Gries , Matthias Hieber , Amru Hussein , Takahito Kashiwabara

Sparse deep neural networks have proven to be efficient for predictive model building in large-scale studies. Although several works have studied theoretical and numerical properties of sparse neural architectures, they have primarily…

Machine Learning · Statistics 2023-09-18 Sanket Jantre , Shrijita Bhattacharya , Tapabrata Maiti

High-dimensional limit theorems have been shown useful to derive tuning rules for finding the optimal scaling in random-walk Metropolis algorithms. The assumptions under which weak convergence results are proved are however restrictive: the…

Methodology · Statistics 2022-02-16 Sebastian M Schmon , Philippe Gagnon

The weak Pareto boundary ($WPB$) refers to a boundary in the objective space of a multi-objective optimization problem, characterized by weak Pareto optimality rather than Pareto optimality. The $WPB$ brings severe challenges to…

Neural and Evolutionary Computing · Computer Science 2026-01-27 Ruihao Zheng , Jingda Deng , Zhenkun Wang
‹ Prev 1 4 5 6 7 8 10 Next ›