English
Related papers

Related papers: Gradient density estimation in arbitrary finite di…

200 papers

Sampling a probability distribution with an unknown normalization constant is a fundamental problem in computational science and engineering. This task may be cast as an optimization problem over all probability measures, and an initial…

Machine Learning · Statistics 2024-09-12 Yifan Chen , Daniel Zhengyu Huang , Jiaoyang Huang , Sebastian Reich , Andrew M. Stuart

Stochastic gradient descent (SGD) is the workhorse of large-scale learning, yet classical analyses rely on assumptions that can be either too strong (bounded variance) or too coarse (uniform noise). The expected smoothness (ES) condition…

Machine Learning · Computer Science 2025-10-28 Yuta Kawamoto , Hideaki Iiduka

We consider the problem of estimating a function $s$ on $[-1,1]^{k}$ for large values of $k$ by looking for some best approximation by composite functions of the form $g\circ u$. Our solution is based on model selection and leads to a very…

Statistics Theory · Mathematics 2013-01-29 Yannick Baraud , Lucien Birgé

We propose a derivative-free trust-region method based on finite-difference gradient approximations for smooth optimization problems with convex constraints. The proposed method does not require computing an approximate stationarity…

Optimization and Control · Mathematics 2025-10-21 Dânâ Davar , Geovani Nunes Grapiglia

Stochastic Gradient Descent (SGD) is a widely deployed optimization procedure throughout data-driven and simulation-driven disciplines, which has drawn a substantial interest in understanding its global behavior across a broad class of…

Optimization and Control · Mathematics 2021-04-02 Vivak Patel , Shushu Zhang

In this paper we analyze the necessary number of samples to estimate the gradient of any multidimensional smooth (possibly non-convex) function in a zero-order stochastic oracle model. In this model, an estimator has access to noisy values…

Machine Learning · Computer Science 2021-07-07 Abdulrahman Alabdulkareem , Jean Honorio

In many areas of applied statistics and machine learning, generating an arbitrary number of independent and identically distributed (i.i.d.) samples from a given distribution is a key task. When the distribution is known only through…

Artificial Intelligence · Computer Science 2021-10-29 Ulysse Marteau-Ferey , Francis Bach , Alessandro Rudi

This paper tackles the challenge of parameter calibration in stochastic models, particularly in scenarios where the likelihood function is unavailable in an analytical form. We introduce a gradient-based simulated parameter estimation…

Machine Learning · Statistics 2025-03-25 Zehao Li , Yijie Peng

We present a real-space formulation for coarse-graining Kohn-Sham Density Functional Theory that significantly speeds up the analysis of material defects without appreciable loss of accuracy. The approximation scheme consists of two steps.…

Computational Physics · Physics 2015-06-11 Phanish Suryanarayana , Kaushik Bhattacharya , Michael Ortiz

Building on the discussion in PRA 93, 042510 (2016), we present a systematic derivation of gradient corrections to the kinetic-energy functional and the one-particle density, in particular for two-dimensional systems. We derive the leading…

Quantum Gases · Physics 2017-09-08 Martin-Isbjörn Trappe , Yink Loong Len , Hui Khoon Ng , Berthold-Georg Englert

When solving finite-sum minimization problems, two common alternatives to stochastic gradient descent (SGD) with theoretical benefits are random reshuffling (SGD-RR) and shuffle-once (SGD-SO), in which functions are sampled in cycles…

Optimization and Control · Mathematics 2022-06-02 Carles Domingo-Enrich

A stationary Gaussian process is said to be long-range dependent (resp., anti-persistent) if its spectral density $f(\lambda)$ can be written as $f(\lambda)=|\lambda|^{-2d}g(|\lambda|)$, where $0<d<1/2$ (resp., $-1/2<d<0$), and $g$ is…

Methodology · Statistics 2012-07-24 Judith Rousseau , Nicolas Chopin , Brunero Liseo

Stochastic gradient (SG) methods are fundamental to system identification and machine learning, enabling online parameter estimation in large-scale and streaming-data settings. As a classical identification method, the SG algorithm has been…

Optimization and Control · Mathematics 2026-05-08 Senhan Yao , Longxu Zhang

In this paper, we develop and analyze sub-sampled trust-region methods for solving finite-sum optimization problems. These methods employ subsampling strategies to approximate the gradient and Hessian of the objective function,…

Optimization and Control · Mathematics 2025-07-24 Max L. N. Goncalves , Geovani N. Grapiglia

Gradient-based methods are well-suited for derivative-free optimization (DFO), where finite-difference (FD) estimates are commonly used as gradient surrogates. Traditional stochastic approximation methods, such as Kiefer-Wolfowitz (KW) and…

Optimization and Control · Mathematics 2025-03-03 Guo Liang , Guangwu Liu , Kun Zhang

We develop new stochastic gradient methods for efficiently solving sparse linear regression in a partial attribute observation setting, where learners are only allowed to observe a fixed number of actively chosen attributes per example at…

Optimization and Control · Mathematics 2018-12-04 Tomoya Murata , Taiji Suzuki

Stochastic approximation (SA) and stochastic gradient descent (SGD) algorithms are work-horses for modern machine learning algorithms. Their constant stepsize variants are preferred in practice due to fast convergence behavior. However,…

Machine Learning · Computer Science 2021-11-12 Zaiwei Chen , Shancong Mou , Siva Theja Maguluri

We consider models of gradient type, which are the densities of a collection of real-valued random variables $\phi :=\{\phi_x: x \in \Lambda\}$ given by $Z^{-1}\exp({-\sum\nolimits_{j \sim k}V(\phi_j-\phi_k)})$. We focus our study on the…

Probability · Mathematics 2019-09-04 Zichun Ye

We wish to compute the gradient of an expectation over a finite or countably infinite sample space having $K \leq \infty$ categories. When $K$ is indeed infinite, or finite but very large, the relevant summation is intractable. Accordingly,…

Machine Learning · Statistics 2019-05-14 Runjing Liu , Jeffrey Regier , Nilesh Tripuraneni , Michael I. Jordan , Jon McAuliffe

For planar and cubic Ising models, we examined two ways of approximation of a spectral density that describes a degeneracy of energy levels. We approximated the exponent of the spectral density by polynomials of even degrees and using our…

Disordered Systems and Neural Networks · Physics 2018-08-15 Leonid Litinskii , Boris Kryzhanovsky
‹ Prev 1 4 5 6 7 8 10 Next ›