English
Related papers

Related papers: Adaptive Convolutions

200 papers

Behavior of neural networks is irremediably determined by the specific loss and data used during training. However it is often desirable to tune the model at inference time based on external factors such as preferences of the user or…

Computer Vision and Pattern Recognition · Computer Science 2023-04-04 Matteo Maggioni , Thomas Tanay , Francesca Babiloni , Steven McDonagh , Aleš Leonardis

The Steklov function $\mu_f(\cdot,t)$ is defined to average a continuous function $f$ at each point of its domain by using a window of size given by $t>0$. It has traditionally been used to approximate $f$ smoothly with small values of $t$.…

Optimization and Control · Mathematics 2020-06-19 Regina S. Burachik , C. Yalçın Kaya

Adjusting for an unmeasured confounder is generally an intractable problem, but in the spatial setting it may be possible under certain conditions. In this paper, we derive necessary conditions on the coherence between the treatment…

Methodology · Statistics 2020-12-23 Yawen Guan , Garritt L. Page , Brian J Reich , Massimo Ventrucci , Shu Yang

Adaptive cubic regularization methods have emerged as a credible alternative to linesearch and trust-region for smooth nonconvex optimization, with optimal complexity amongst second-order methods. Here we consider a general/new class of…

Optimization and Control · Mathematics 2018-11-20 Coralia Cartis , Nicholas I. M. Gould , Philippe L. Toint

Convolutional neural network is an important model in deep learning. To avoid exploding/vanishing gradient problems and to improve the generalizability of a neural network, it is desirable to have a convolution operation that nearly…

Machine Learning · Computer Science 2019-06-25 Peichang Guo , Qiang Ye

The free metaplectic transformation (FMT) is widely used in many fields such as filter design, pattern recognition, image processing and optics. In order to obtain a more concise and intuitive convolution form, this paper studies two kinds…

General Mathematics · Mathematics 2022-06-28 Hui Zhao , Bing-Zhao Li

Many neural network architectures rely on the choice of the activation function for each hidden layer. Given the activation function, the neural network is trained over the bias and the weight parameters. The bias catches the center of the…

Machine Learning · Computer Science 2019-10-01 Farnoush Farhadi , Vahid Partovi Nia , Andrea Lodi

We focus on establishing the foundational paradigm of a novel optimization theory based on convolution with convex kernels. Our goal is to devise a morally deterministic model of locating the global optima of an arbitrary function, which is…

Optimization and Control · Mathematics 2025-03-31 Zhipeng Lu

It is well-known that the high computational complexity and the insufficient samples in large-scale array signal processing restrict the real-world applications of the conventional full-dimensional adaptive beamforming (sample matrix…

Information Theory · Computer Science 2014-05-20 Hu Xie , Da-Zheng Feng , Ming-Dong Yuan

Smooth activation functions are ubiquitous in modern deep learning, yet their theoretical advantages over non-smooth counterparts remain poorly understood. In this work, we study both approximation and statistical properties of neural…

Machine Learning · Statistics 2026-03-03 Yuhao Liu , Zilin Wang , Lei Wu , Shaobo Zhang

We tensorize the Faber spline system from [14] to prove sequence space isomorphisms for multivariate function spaces with higher mixed regularity. The respective basis coefficients are local linear combinations of discrete function values…

Functional Analysis · Mathematics 2020-04-08 Nadiia Derevianko , Tino Ullrich

The approximation properties of the finite element method can often be substantially improved by choosing smooth high-order basis functions. It is extremely difficult to devise such basis functions for partitions consisting of arbitrarily…

Numerical Analysis · Mathematics 2021-01-18 Eky Febrianto , Michael Ortiz , Fehmi Cirak

Fine-tuning is widely applied in image classification tasks as a transfer learning approach. It re-uses the knowledge from a source task to learn and obtain a high performance in target tasks. Fine-tuning is able to alleviate the challenge…

Computer Vision and Pattern Recognition · Computer Science 2022-07-27 Xuyang Shen , Jo Plested , Sabrina Caldwell , Yiran Zhong , Tom Gedeon

We study the problem of optimizing a function under a \emph{budgeted number of evaluations}. We only assume that the function is \emph{locally} smooth around one of its global optima. The difficulty of optimization is measured in terms of…

Machine Learning · Computer Science 2019-02-26 Peter L. Bartlett , Victor Gabillon , Michal Valko

The mutation strength adaptation properties of a multi-recombinative $(\mu/\mu_I, \lambda)$-ES are studied for isotropic mutations. To this end, standard implementations of cumulative step-size adaptation (CSA) and mutative self-adaptation…

Neural and Evolutionary Computing · Computer Science 2024-08-20 Amir Omeradzic , Hans-Georg Beyer

In the present paper we consider the problem of estimating a periodic $(r+1)$-dimensional function $f$ based on observations from its noisy convolution. We construct a wavelet estimator of $f$, derive minimax lower bounds for the $L^2$-risk…

Statistics Theory · Mathematics 2013-05-24 Rida Benhaddou , Marianna Pensky , Dominique Picard

We consider the estimation of the slope function in functional linear regression, where scalar responses are modeled in dependence of random functions. Cardot and Johannes [J. Multivariate Anal. 101 (2010) 395-408] have shown that a…

Statistics Theory · Mathematics 2013-02-19 Fabienne Comte , Jan Johannes

We study the convolution function $$ C[f(x)] := \int_1^x f(y)f({x\over y}) {{\rm d} y\over y} $$ when $f(x)$ is a suitable number-theoretic error term. Asymptotics and upper bounds for $C[f(x)]$ are derived from mean square bounds for…

Number Theory · Mathematics 2010-11-03 Aleksandar Ivic

We present here a new method for approximating functions defined on superreflexive Banach spaces by differentiable functions with $\alpha$-H\"older derivatives (for some $0<\alpha\leq 1$). The smooth approximation is given by means of an…

Functional Analysis · Mathematics 2016-09-07 Manuel Cepedello Boiso

This paper deals with estimation with functional covariates. More precisely, we aim at estimating the regression function $m$ of a continuous outcome $Y$ against a standard Wiener coprocess $W$. Following Cadre and Truquet (2015) and Cadre,…

Statistics Theory · Mathematics 2020-11-23 Karine Bertin , Nicolas Klutchnikoff