English
Related papers

Related papers: An improved example for an autoconvolution inequal…

200 papers

In neural network training, RMSProp and Adam remain widely favoured optimisation algorithms. One of the keys to their performance lies in selecting the correct step size, which can significantly influence their effectiveness. Additionally,…

Machine Learning · Computer Science 2024-04-05 Alokendu Mazumder , Rishabh Sabharwal , Manan Tayal , Bhartendu Kumar , Punit Rathore

In this paper, the optimal convergence rate $O\left(N^{-1/2}\right)$ (where $N$ is the total number of iterations performed by the algorithm), without the presence of a logarithmic factor, is proved for mirror descent algorithms with…

Optimization and Control · Mathematics 2025-06-04 Mohammad Alkousa , Fedor Stonyakin , Asmaa Abdo , Mohammad Alcheikh

We study the asymptotic convergence of solutions as $t\rightarrow\infty$ of $\partial_t u=-f(u)+\int f(u)$, a nonlocal differential equation that is formally a gradient flow in a constant-mass subspace of $L^2$ arising from simplified…

Classical Analysis and ODEs · Mathematics 2024-09-16 Sangmin Park , Robert L. Pego

This paper investigates some aspects of the variational behaviour of nonsmooth functions, with special emphasis on certain stability phenomena. Relationships linking such properties as sharp minimality, superstability, error bound and…

Optimization and Control · Mathematics 2014-10-10 Amos Uderzo

We consider optimizing a function smooth convex function $f$ that is the average of a set of differentiable functions $f_i$, under the assumption considered by Solodov [1998] and Tseng [1998] that the norm of each gradient $f_i'$ is bounded…

Optimization and Control · Mathematics 2013-08-30 Mark Schmidt , Nicolas Le Roux

Constant step-size Stochastic Gradient Descent exhibits two phases: a transient phase during which iterates make fast progress towards the optimum, followed by a stationary phase during which iterates oscillate around the optimal point. In…

Machine Learning · Computer Science 2020-07-02 Scott Pesme , Aymeric Dieuleveut , Nicolas Flammarion

Let $f\in L_{2\pi}$ be a real-valued even function with its Fourier series $ \frac{a_{0}}{2}+\sum_{n=1}^{\infty}a_{n}\cos nx,$ and let $S_{n}(f,x), n\geq 1,$ be the $n$-th partial sum of the Fourier series. It is well-known that if the…

Classical Analysis and ODEs · Mathematics 2007-05-23 Dan Sheng Yu , Ping Zhou , Song Ping Zhou

The Riesz-Sobolev inequality relates the convolution of nonnegative functions on Euclidean space to the convolution of their symmetric nonincreasing rearrangements. We show that for dimension one, for indicator functions of sets, if the…

Classical Analysis and ODEs · Mathematics 2011-12-19 Michael Christ

We provide non-asymptotic bounds for the well-known temporal difference learning algorithm TD(0) with linear function approximators. These include high-probability bounds as well as bounds in expectation. Our analysis suggests that a…

Machine Learning · Computer Science 2015-09-02 Nathaniel Korda , L. A. Prashanth

We present here a new method for approximating functions defined on superreflexive Banach spaces by differentiable functions with $\alpha$-H\"older derivatives (for some $0<\alpha\leq 1$). The smooth approximation is given by means of an…

Functional Analysis · Mathematics 2016-09-07 Manuel Cepedello Boiso

We show how to obtain improved active learning methods in the agnostic (adversarial noise) setting by combining marginal leverage score sampling with non-independent sampling strategies that promote spatial coverage. In particular, we…

Machine Learning · Computer Science 2024-05-07 Atsushi Shimizu , Xiaoou Cheng , Christopher Musco , Jonathan Weare

Convolutional autoencoders have emerged as popular methods for unsupervised defect segmentation on image data. Most commonly, this task is performed by thresholding a pixel-wise reconstruction error based on an $\ell^p$ distance. This…

Computer Vision and Pattern Recognition · Computer Science 2019-04-09 Paul Bergmann , Sindy Löwe , Michael Fauser , David Sattlegger , Carsten Steger

Let $\Omega \subset \mathbb{R}^n$ be a convex domain and let $f:\Omega \rightarrow \mathbb{R}$ be a positive, subharmonic function (i.e. $\Delta f \geq 0$). Then $$ \frac{1}{|\Omega|} \int_{\Omega}{f dx} \leq \frac{c_n}{ |\partial \Omega| }…

Structured constraints in Machine Learning have recently brought the Frank-Wolfe (FW) family of algorithms back in the spotlight. While the classical FW algorithm has poor local convergence properties, the Away-steps and Pairwise FW…

Optimization and Control · Mathematics 2022-09-09 Fabian Pedregosa , Geoffrey Negiar , Armin Askari , Martin Jaggi

Current state-of-art feature-engineered and end-to-end Automated Essay Score (AES) methods are proven to be unable to detect adversarial samples, e.g. the essays composed of permuted sentences and the prompt-irrelevant essays. Focusing on…

Computation and Language · Computer Science 2019-12-23 Jiawei Liu , Yang Xu , Yaguang Zhu

The Bregman distance is a central tool in convex optimization, particularly in first-order gradient descent and proximal-based algorithms. Such methods enable optimization of functions without Lipschitz continuous gradients by leveraging…

Optimization and Control · Mathematics 2025-04-28 Max Nilsson , Pontus Giselsson

We consider the model of nonregular nonparametric regression where smoothness constraints are imposed on the regression function $f$ and the regression errors are assumed to decay with some sharpness level at their endpoints. The aim of…

Statistics Theory · Mathematics 2014-10-02 Moritz Jirak , Alexander Meister , Markus Reiß

We analyze the constant step size subgradient method on nonsmooth, nonconvex functions. We identify geometric assumptions on the objective function under which i) its domain admits a partition (stratification) into smooth manifolds (strata)…

Optimization and Control · Mathematics 2026-04-21 Evgenii Chzhen , Sholom Schechtman

This paper develops tests for inequality constraints of nonparametric regression functions. The test statistics involve a one-sided version of $L_p$-type functionals of kernel estimators $(1 \leq p < \infty)$. Drawing on the approach of…

Statistics Theory · Mathematics 2023-08-28 Sokbae Lee , Kyungchul Song , Yoon-Jae Whang

In this paper, we propose a simple, fast and easy to implement algorithm LOSSGRAD (locally optimal step-size in gradient descent), which automatically modifies the step-size in gradient descent during neural networks training. Given a…

Machine Learning · Computer Science 2019-11-26 Bartosz Wójcik , Łukasz Maziarka , Jacek Tabor